cd/entity/BM25· home› entities› BM25
grep -l @bm25 /news/*.json | wc -l → 127

BM25

mentions 127 type Organization page 3/7 feed RSS

// recent coverage 127 mentions

19:36
2026-08-22
dev.to
artificial-intelligence

9 RAG Techniques That Actually Improve Retrieval Quality

A developer's guide details nine techniques to improve retrieval quality in Retrieval-Augmented Generation (RAG) systems, including reranking, hybrid search, and context compression. The techniques ad…

05:00
2026-08-22
dev.to
artificial-intelligence

Simple BM25 outperforms agents on large corpora

A new study finds that lexical BM25 retrieval outperforms sophisticated agentic search methods on large corpora, achieving 50.5 accuracy versus 30.7 for File-System Agent at the largest scale. The gap…

15:01
2026-08-21
pub.towardsai.net
artificial-intelligence

LAI #139: Fewer Tokens Cost Us More

A production AI tutor's context engineering experiments, detailed in the newsletter LAI #139, found that reducing tokens via summarization increased costs by roughly 2x despite sending 41% fewer token…

04:00
2026-08-20
machinebrief.com
natural-language-processing

GreekBarRetrieval: A Benchmark for Greek Statutory Retrieval

GreekBarRetrieval, a new public benchmark for Greek statutory retrieval derived from GreekBarBench, comprises 283 bar-exam questions and 6,308 candidate statutory articles. Testing three BM25 variants…

05:00
2026-08-16
dev.to
artificial-intelligence

AI/ML Research Digest — Aug 02, 2026

New research highlights advances in AI reliability and efficiency. Σ-Mem and LedgerMind introduce explicit reliability modeling and structured evidence ledgers to reduce hallucinations and enable prov…

19:28
2026-08-14
aletheionagi.com
ai-agents

Show HN: AletheionAGI – Grounding enforcement for AI agents

AletheionAGI, a new grounding enforcement layer for AI agents, claims to block unsupported claims before they reach customers, reporting 93.6% Recall@5 and a 66.5% diagnostic answer score on 979 froze…

09:35
2026-08-10
dev.to
machine-learning

BM25 Length Normalization: Why Long RAG Chunks Never Rank

A developer found that BM25's default length normalization parameter (b=0.75) systematically penalizes long RAG chunks, causing exact-answer chunks to rank below irrelevant headings. The analysis show…

22:27
2026-08-04
marktechpost.com
artificial-intelligence

Pixel-Native RAG: A Practical Guide to Visual Document Indexing

A new tutorial from StarTrail-org introduces PixelRAG, a pixel-native retrieval-augmented generation pipeline that renders web pages and PDFs as images, divides them into overlapping tiles, and genera…

10:35
2026-08-04
discuss.huggingface.co
artificial-intelligence

Chunking strategy for governments and internal org docs

A developer seeking advice on building a retrieval-augmented generation (RAG) system for government and internal organizational documents asks about optimal chunk sizes for BM25 and semantic search (c…

← prev page 3 / 7 next →
// co-occurs with top 8 entities
// topics top 6 topics