cd /news/artificial-intelligence/hc-rag-evidence-centric-retrieval-au… · home topics artificial-intelligence article
[ARTICLE · art-96289] src=arxiv.org ↗ pub= topic=artificial-intelligence verified=true sentiment=↑ positive

HC-RAG: Evidence-Centric Retrieval-Augmented Generation over Heterogeneous Financial Filings

Researchers propose HC-RAG, a hierarchical cross-modal retrieval-augmented generation framework for evidence-centric financial question answering, which organizes filings into a typed financial evidence graph and routes evidence by query intent. On the new Multi-Doc-2025 benchmark of 2,327 expert-verified QA pairs from 179 SEC 10-K filings, HC-RAG outperforms GraphRAG by 10.9 F1 points, and it beats RAPTOR by 6.6 F1 points on DocFinQA.

read1 min views1 publishedAug 14, 2026

arXiv:2608.12335v1 Announce Type: new Abstract: Financial question answering over annual reports requires more than retrieving semantically similar passages. It often involves identifying relevant companies and fiscal years, locating standardized filing sections, collecting textual and tabular evidence, and checking answers against the original documents. Existing RAG systems, however, usually flatten long filings into unordered chunks, pay limited attention to the typed structure of financial reports, and use fixed text-table fusion strategies without considering query intent. To address these limitations, we propose \textbf{HC-RAG}, a hierarchical cross-modal retrieval-augmented generation framework for evidence-centric financial QA. HC-RAG organizes filings into a typed financial evidence graph with documents, sections, text units, table units, and metadata nodes. It retrieves evidence through document-section-unit paths, aligns textual and tabular evidence in a shared retrieval space, and routes evidence according to four semantic intents: calculation, trend, fact, and comparison. We further introduce \textbf{Multi-Doc-2025}, a benchmark containing 2,327 expert-verified QA pairs from 179 SEC 10-K filings of 87 S&P 500 companies across fiscal years 2022--2024, with labels for intent, difficulty, and structural evidence attributes. Experiments on public financial QA benchmarks and Multi-Doc-2025 show that HC-RAG improves both answer quality and evidence localization, especially in long-document, table-related, and cross-document settings. HC-RAG outperforms RAPTOR by 6.6 F1 points on DocFinQA and GraphRAG by 10.9 F1 points on Multi-Doc-2025. Evidence-level analysis and ablation studies show that the improvements mainly come from more accurate section localization, table grounding, cross-document evidence aggregation, and intent-aware text-table routing.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @hc-rag 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/hc-rag-evidence-cent…] indexed:0 read:1min 2026-08-14 ·