Hugging Face Blog

Building Deep Research: How we Achieved State of the Art

OpenAI Blog
Feb 2, 2025

Introducing deep research

An agent that uses reasoning to synthesize large amounts of online information and complete multi-step research tasks for you. Available to Pro users today, Plus and Team next.

OpenAI Blog
Feb 25, 2025

Deep research System Card

This report outlines the safety work carried out prior to releasing deep research including external red teaming, frontier risk evaluations according to our Preparedness Framework, and an overview of the mitigations we built in to address key risk areas.

arXiv AI
6d ago

LongCat-DeepResearch Technical Report

LongCat-DeepResearch is a deep research system that merges an enhanced LongCat model with a multi‑agent workflow to produce comprehensive, evidence‑grounded reports. The workflow separates global planning from detailed investigation, using planning agents to create a ResearchSpec and research agents to draft sections in parallel, followed by targeted local revisions guided by global review. The system achieves strong benchmark scores, including 55.25 on DeepResearchBench and 79.83 on ResearchRubrics, and shows benefits from combining planning perspectives and additional editing for readability.

By Meituan LongCat Team, He Zhu, Yue Xu, Wanli Wu, Haolin Ren, Yuxin Bian, Jiarui Zhao, Rongzhi Zhang, Quanchi Weng, Jinghao Cui, Yu Fan, Yuhan Liu, Yunhu Ye, Jiyuan Ren, Fengcheng Yuan, Zhao Yang, Jiacheng Zhang, Yuchuan Dai, Ruixuan Xiao, Haozhe Sun, Xiangyuan Liu, Cheng Sun, Yao Du, Yiming Hao, Hongbo Guo, Shuo He, Lei Wang, Xunliang Cai, Yan Chen, Fan Yang, Lingchuan Liu
arXiv Computation and Language
Aug 27

Mind2Report: Expert-Level Commercial Report Synthesis via Cognitive Deep Research Agent

Mind2Report is a cognitive deep research agent designed to produce expert-level commercial reports from large, noisy web sources. It first clarifies detailed commercial intent to build a structured outline, then recursively gathers and validates evidence into a research memory that evolves with the outline, enabling iterative synthesis of comprehensive reports. The authors also introduce QRC‑Eval, a benchmark of 200 real-world commercial tasks, and show through extensive experiments that Mind2Report outperforms existing proprietary and open-source deep research agents, with ablation studies confirming the contribution of each component.

By Mingyue Cheng, Daoyu Wang, Qi Liu, Shuo Yu, Xiaoyu Tao, Yuqian Wang, Chengzhong Chu, Yu Duan, Mingkang Long, Enhong Chen
OpenAI Blog
Aug 29, 2016

Infrastructure for deep learning

Deep learning is an empirical science, and the quality of a group’s infrastructure is a multiplier on progress. Fortunately, today’s open-source ecosystem makes it possible for anyone to build great deep learning infrastructure.