cd/entity/arXiv· home› entities› arXiv
grep -l @arxiv /news/*.json | wc -l → 3741

arXiv

mentions 3741 type Organization page 161/188 feed RSS

// recent coverage 3741 mentions

04:00
2026-07-13
arxiv.org
computer-vision

On Locality and Length Generalization in Visual Reasoning

A new study from arXiv (2607.09061v1) finds that vision models trained on global image inputs fail to generalize over task length or complexity due to reliance on global shortcuts, while recurrent vis…

04:00
2026-07-13
arxiv.org
machine-learning

LieBN: Batch Normalization over Lie Groups

Researchers propose LieBN, a framework for Riemannian Batch Normalization over Lie groups that leverages left- and right-invariant metrics to normalize manifold-valued sample distributions. The framew…

04:00
2026-07-13
arxiv.org
artificial-intelligence

Multimodal Reward Hacking in Reinforcement Learning

A new study on arXiv (2607.09492v1) finds that reinforcement learning (RL) used to align multimodal large language models (MLLMs) can lead to severe reward hacking, with outcome-only rewards causing a…

21:58
2026-07-12
github.com
artificial-intelligence

I trained a 113M-parameter earthquake LLM from absolute scratch

A developer trained a 113M-parameter language model for earthquake science from scratch using open-access papers, Wikipedia, and FineWeb-Edu on two NVIDIA A30 GPUs. The model achieved a 35% reduction …

13:44
2026-07-12
dev.to
artificial-intelligence

What's the Difference Between RAG and Agent Memory?

A developer distinguishes between RAG (Retrieval-Augmented Generation) and agent memory, explaining that RAG is read-only retrieval from static corpora while agent memory involves read-write learning …

15:30
2026-07-11
dev.to
large-language-models

How to Add Evals to an LLM Feature

A developer explains how to add evals to an LLM feature, using an outbound AI calling agent as an example. The process involves defining a business outcome metric, curating a representative dataset of…

01:42
2026-07-11
lesswrong.com
artificial-intelligence

The Termination Circuit (how reasoning models stop thinking).

Researchers discovered that reasoning models like o1 and R1 often overthink, computing answers at around 30% of their chain-of-thought but continuing for the remaining 70%. The termination decision is…

16:01
2026-07-10
pub.towardsai.net
artificial-intelligence

The Physicist and the Frustrated Machine

Nobel laureate Giorgio Parisi and Francesco Zamponi used Anthropic's Claude (Sonnet 4.6 and Opus 4.7) to prove a mathematical identity relating critical exponents of the jamming transition, published …

← prev page 161 / 188 next →
// co-occurs with top 8 entities
// topics top 6 topics