cd/entity/arXiv· home› entities› arXiv
grep -l @arxiv /news/*.json | wc -l → 3740

arXiv

mentions 3740 type Organization page 139/187 feed RSS

// recent coverage 3740 mentions

18:32
2026-07-31
arxiv.org
artificial-intelligence

Orca-Bench: How Ready Are Language Model Agents for Oncall?

A new benchmark, ORCA-bench, shows that frontier language model agents achieve only 25.3% root cause analysis accuracy on Medium-difficulty oncall tasks and 10.0% on Hard tasks, with the best performa…

18:10
2026-07-31
arxiv.org
artificial-intelligence

Neuro-Inspired Inverse Learning for Planning and Control

Researchers led by Tonio Ball introduced the Inverter framework, a neuro-inspired approach for embodied planning and control that uses paired forward/inverse internal models and open-loop multi-step m…

17:27
2026-07-31
lesswrong.com
artificial-intelligence

How to Measure Intelligence Beyond Human Scale?

Researchers propose 'adversarial psychometrics' to measure AI intelligence beyond human scale, where participants generate questions and are rewarded for separating each other's capabilities without a…

09:12
2026-07-31
aiproductopportunity.com
ai-startups

Show HN: AI Product Opportunity

AI Product Opportunity, a new platform showcased on Hacker News, claims to help founders discover and validate AI product opportunities using evidence from 15 sources, tracking 3,000+ opportunities, 1…

04:00
2026-07-31
arxiv.org
artificial-intelligence

MeshFM: 2D Features Are All You Need for 3D Shape Understanding

Researchers introduced MeshFM, a feedforward framework that distills 2D features from visual foundation models into 3D, enabling rich 3D shape understanding without 3D annotations. The method, which u…

04:00
2026-07-31
arxiv.org
artificial-intelligence

Position: Evaluation Scores Are Perishable Knowledge Claims

A new arXiv paper (2607.26191v1) argues that language model evaluation scores should be treated as perishable epistemic claims with formal metadata, warning that averaging signals from automated metri…

04:00
2026-07-31
arxiv.org
artificial-intelligence

Evidence-Ledger Adjudication for Claim-Evidence Traceability

A new arXiv paper (2607.26512v1) introduces evidence-ledger adjudication, a workflow that pairs AI-generated claims with evidence packets and routes unsupported or contradicted claims back to authors.…

← prev page 139 / 187 next →
// co-occurs with top 8 entities
// topics top 6 topics