When Search Agents Should Ask: DiscoBench for Clarification-Aware Deep Search

wpnews.pro

cd /news/large-language-models/when-search-agents-should-ask-discob… · home › topics › large-language-models › article

[ARTICLE · art-42912] src=arxiv.org ↗ pub=2026-06-29T04:00Z topic=large-language-models verified=true sentiment=· neutral

When Search Agents Should Ask: DiscoBench for Clarification-Aware Deep Search

Researchers introduced DiscoBench, a benchmark for clarification-aware deep search, to evaluate whether LLM-powered search agents can identify ambiguity, ask effective questions, and recover correct reasoning paths. The benchmark contains 211 samples across 11 domains, and experiments show that current agents often fail to ask for clarification, performing worse than direct guessing.

read1 min views1 publishedJun 29, 2026

arXiv:2606.27669v1 Announce Type: new Abstract: Search agents powered by large language models (LLMs) are increasingly used to solve complex information-seeking tasks, requiring multi-step retrieval and reasoning to fulfill user goals. However, existing benchmarks often assume that user queries are complete and explicit, overlooking the fact that real-world search requests are frequently vague, underspecified, or even factually incorrect. In deep search scenarios, such ambiguity can propagate along multi-step reasoning chains and lead agents toward incorrect search trajectories. To address this gap, we introduce DiscoBench, a benchmark for clarification-aware deep search, designed to evaluate whether search agents can proactively identify ambiguity, ask effective clarification questions, and recover correct reasoning paths through user interaction. DiscoBench contains 211 samples and 463 ambiguity instances across 11 real-world domains, covering four ambiguity types. We further design a user simulator for multi-turn interaction and evaluate model performance from four perspectives: task utility, ambiguity detection, interaction strategy, and cost efficiency. Experiments on representative LLMs show that ambiguity detection and effective clarification are distinct capabilities, and that repeatedly searching instead of asking for clarification often performs worse than direct guessing, highlighting a critical gap between retrieval ability and interactive problem-solving in current search agents.

source & further reading

arxiv.org — original article

~/api · this article 200

$curl api.wpnews.pro/v1/news/when-search-agents-shoul…

Read original on arxiv.org → arxiv.org/abs/2606.27669

mentioned entities

DiscoBench

arXiv

metadata

slugwhen-search-agents-should-ask-discobench-for-clarification-aware-deep-search

topic#large-language-models

secondary3 topics

sentimentneutral

canonicalarxiv.org

navigation

← prevv0.5.6

── more in #large-language-models 4 stories · sorted by recency

arxiv.org · 29 Jun · #large-language-models

Ko-WideSearch: A Korean Breadth-Search Benchmark for Exhaustive Set Enumeration by Web Agents

arxiv.org · 29 Jun · #large-language-models

Supersede: Diagnosing and Training the Memory-Update Gap in LLM Agents

arxiv.org · 29 Jun · #large-language-models

Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning

arxiv.org · 29 Jun · #large-language-models

Developmental approach reveals the statistical learning of Neural Language Models: Transformers generalize from the most abstract statistical patterns

── more on @discobench 3 stories trending now

wpnews · 28 May · #ai-startups

[AINews] Cognition raises $1B in $26B Series D

wpnews · 5 Jun · #ai-agents

Miasma Worm Targets AI Coding Agents via GitHub Repos

wpnews · 28 Jun · #ai-agents

OpenCode v1.17: Session Snapshots Undo Your AI Agent

sponsored brought to you by zahid.host 4,200+ EU-deployed projects

reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main

→ Live at https://your-agent.zahid.host ✓

Get free account → Pricing

from €0/mo · no card required