#DeepSearchQA
<a href="https://gigazine.net/news/20251212-gemini-deep-research/" class="hover:underline text-blue-600 dark:text-sky-400 no-card-link" target="_blank" rel="noopener" data-link="bsky">gigazine.net/news/20251212...
Googleが「Gemini Deep Research エージェント」をリリース&ベンチマーク「DeepSearchQA」オープンソース化
Googleが「Gemini Deep Research エージェント」をリリース&ベンチマーク「DeepSearchQA」オープンソース化
gigazine.net
December 14, 2025 at 8:50 AM
comparing to gpt-oss:120b, looks like shimmer is the better agentic model, but gpt-oss is better at traditional LLM tasks
August 10, 2026 at 2:15 PM
🚀 New on Kaggle Benchmarks: DeepSearchQA developed by Google DeepMind!

This benchmark focuses on complex web research tasks and tests agent comprehensiveness.

Check the leaderboard: www.kaggle.com/benchmarks/g...
December 11, 2025 at 6:30 PM
Google DeepMind, geliştiricilerin yeni Interactions API üzerinden erişebileceği geliştirilmiş Cem Ceminay (Gemini) Deep Research ajanı ve yeni DeepSearchQA benchmarkını yayınladı.
Build with Gemini Deep Research
We have reimagined Gemini Deep Research to be more powerful than ever, now accessible to developers via the new Interactions API.
blog.google
December 12, 2025 at 8:56 AM
#Google launched Deep Research and Deep Research Max: autonomous agents that blend open web and private data via MCP in one API call.

Built on Gemini 3.1 Pro. Max hits 93.3% on DeepSearchQA.

Caveats: API only, 60-min limit, MCP quality depends on the server, benchmarks ≠ real research.
Deep Research Max: a step change for autonomous research agents
Introducing Deep Research and Deep Research Max, the next generation of Google’s autonomous research agents.
blog.google
April 22, 2026 at 9:19 PM
🥝 Meet Kimi K2.5, Open-Source Visual Agentic Intelligence. (1/4)
January 27, 2026 at 6:48 PM
Google launched Deep Research and Deep Research Max, Gemini 3.1 Pro-powered AI agents that combine web search and proprietary enterprise data via the Model Context Protocol, generate native charts, and achieved 93.3% on DeepSearchQA while targeting finance and life sciences.
Google’s new Deep Research and Deep Research Max agents can search the web and your private data
Google unveiled Deep Research and Deep Research Max, new Gemini 3.1 Pro-powered AI agents that combine web search, proprietary enterprise data, MCP integrations, and native charts to automate high-stakes research workflows in finance, life sciences, and market intelligence.
venturebeat.com
April 21, 2026 at 11:22 PM
The harness, not the model, may be the lever. JIT-Agent synthesizes task-specific agent scaffolds on the fly: DeepSeek-V4-Flash passes GPT-5.6 by 9.1 points on DeepSearchQA and 4.3 on OdysseyBench, while GLM-5.2 gains up to 20.2. Scaffolding is becoming a trainable layer.
August 28, 2026 at 2:29 AM
Google launches autonomous research agents for finance, life sciences and market intelligence
Google launches autonomous research agents for finance, life sciences and market intelligence
Deep Research Max scores 93.3% on DeepSearchQA, up from 66.1% in December, as the company targets enterprise workflows
dlvr.it
April 22, 2026 at 4:37 PM
Google has also open-sourced DeepSearchQA, a new benchmark for evaluating multi-step research agents.
December 15, 2025 at 8:26 PM
Google Releases a "More Powerful" Deep Research Agent & Open Sources a New Web Research Agent Benchmark, #DeepSearchQA blog.google/innovation-a... & Gemini 3 Flash Comes to Gemini App blog.google/products-and... #AI #LLMs #GenAI #deepresearch
January 9, 2026 at 2:03 PM
- 구글이 Interactions API를 통해 제미나이 딥 리서치를 출시하여 개발자가 고급 자율 연구 기능을 애플리케이션에 임베드할 수 있게 함

- 제미나이 3 프로 기반으로 반복적으로 조사를 계획하고 쿼리를 생성하며 지식 격차를 식별함

- HLE 에서 46.4%, DeepSearchQA 벤치마크에서 66.1% 달성

오늘 보니 제미나이 공웹 빠른 모드로 하니 딥리서치 생기더군요. (원래 있었던 것 같긴합니다만)

Build with Gemini Deep Research

blog.google/technology/d...
Build with Gemini Deep Research
We have reimagined Gemini Deep Research to be more powerful than ever, now accessible to developers via the new Interactions API.
blog.google
December 12, 2025 at 12:38 PM
Kimi正式开源旗舰模型K2.6,在代码能力、长程任务执行和Agent集群协作上实现重大突破,核心亮点如下:

1. 性能全面领先
•基准测试:在博士级考试Humanity’s Last Exam(54.0%)、Agent检索DeepSearchQA(92.5%)、软件工程SWE-Bench Pro(58.6%)等评测中超越GPT-5.4、Gemini 3.1 Pro等主流模型。
•短板:多语言编程(SWE-bench)、复杂工具调度(Toolathlon)及视觉任务(MathVision)仍小幅落后顶尖模型
April 21, 2026 at 8:57 AM
Google DeepMind launches an enhanced Gemini Deep Research agent accessible to developers via its new Interactions API, along with a new DeepSearchQA benchmark (The Keyword)

Main Link | Techmeme Permalink
December 11, 2025 at 5:40 PM
Gemini’s Deep Research agent just aced Humanity’s Last Exam, topping HLE, DeepSearchQA and leading BrowseComp. Curious how it stacks up against Google Search and NotebookLM? Dive into the benchmark details! #GeminiDeepResearch #DeepSearchQA #BrowseComp

🔗 aidailypost.com/news/gemini-...
December 11, 2025 at 6:42 PM
Google lancia Gemini Deep Research 3 Pro, agente di ricerca autonomo per app via API.
Navigazione migliorata, meno allucinazioni, punteggi 46,4 % su HumanityExam, 66,1 % DeepSearchQA, 59,2 % BrowseComp 🚀🤖#gemini #ai #ricerca
December 12, 2025 at 11:09 AM
This tweet appeared under this Techmeme headline:

@google:

We're also open-sourcing DeepSearchQA, a new benchmark to evaluate agents on complex search tasks. Learn more about these updates on the Keyword Blog ↓ https://blog.google/...
December 11, 2025 at 7:52 PM
Kimi K2.6 released beating closed models on multiple benchmarks with SOTA scores across HLE, DeepSearchQA and SWE-Bench Pro.
April 21, 2026 at 11:32 AM
🤖 #AgentSwarm now scales to 300 parallel sub-agents coordinating 4,000 steps simultaneously — up from 100 agents in K2.5

📊 80.2% on SWE-Bench Verified, 54.0% on HLE-Full with tools (beats GPT-5.4), 92.5 F1 on DeepSearchQA #coding #MachineLearning
April 21, 2026 at 12:04 AM