cd/entity/HotpotQA· home entities HotpotQA
grep -l @hotpotqa /news/*.json | wc -l → 27

HotpotQA

mentions 27 type Organization page 1/2 feed RSS

// recent coverage 27 mentions

04:00
2026-08-11
machinebrief.com
artificial-intelligence

SAGE: SLO-Aware Adaptive Retrieval for Production RAG Systems

A new paper on arXiv (2608.08237v1) proposes SAGE, a learned SLO-aware adaptive retrieval policy for production RAG systems that dynamically selects the number of passages per query. On Natural Questi…

17:00
2026-08-07
usewire.io
artificial-intelligence

Agentic RAG fails before the reasoning starts

A 2026 study by Daeyoung Roh and Donghee Han of 12,000 paired agent trajectories across HotpotQA, 2WikiMultiHopQA, and MuSiQue found that agentic RAG failures often stem from procedural errors: on MuS…

06:16
2026-08-05
pub.towardsai.net
large-language-models

Your Hallucination Benchmark Is Measuring Your Detector

A study labeling 7,440 answers from four open-weight LLMs found that more than half of the hallucination labels were incorrect, and correcting them reordered the results. The author, Priyanshi Jain, r…

06:40
2026-07-15
machinebrief.com
artificial-intelligence

Evidence Selection in RAG with QUBO

A new approach using Quadratic Unconstrained Binary Optimization (QUBO) for evidence selection in retrieval-augmented generation (RAG) systems achieves competitive exact-match and token-F1 performance…

06:37
2026-07-13
machinebrief.com
artificial-intelligence

Memory-Managed Attention: Redefining AI's Long-Term Memory

A new study on memory-managed long-context attention achieves a perfect 1.000 score on Track A, a controlled retrieval task, compared to a 0.333 baseline, and outperforms dense retrieval methods by up…

07:27
2026-07-10
machinebrief.com
artificial-intelligence

DeepSearch-Evolve: The Next Step in Self-Improving AI Agents

DeepSearch-Evolve introduces a self-distillation framework for training web agents in a controlled environment, achieving state-of-the-art results on benchmarks like BrowseComp, GAIA, and HotpotQA wit…

04:00
2026-07-01
arxiv.org
large-language-models

Contrastive Reflection for Iterative Prompt Optimization

Researchers introduced Contrastive Reflection, an iterative prompt-optimization framework for agentic information retrieval workflows, which uses error-anchored behavioral slices and contrastive examp…

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics