cd/entity/arXiv· home› entities› arXiv
grep -l @arxiv /news/*.json | wc -l → 3738

arXiv

mentions 3738 type Organization page 131/187 feed RSS

// recent coverage 3738 mentions

01:07
2026-08-10
arxiv.org
large-language-models

LLMs Get Lost in Evolving User Intent

A study submitted to arXiv on 22 Jul 2026 found that large language models (LLMs) suffer substantial performance drops when user intent evolves across multi-turn conversations, despite strong performa…

22:06
2026-08-09
letsdatascience.com
artificial-intelligence

Overthinking Method Exposes Hidden Behaviors in Qwen3-VL Tests

An ICML 2026 paper submitted July 9 reports that amplifying the weight-space difference between Qwen3-VL instruct and reasoning checkpoints made planted secrets or unintended behaviors surface up to 1…

20:48
2026-08-06
aiunderstanding.org
artificial-intelligence

Study Finds Frontier AI Models Split Under Steering Pressure

A preprint posted on August 6 found that six frontier language models—Claude Opus 4.7, GPT-5, Gemini 2.5 Pro, DeepSeek-R1, Qwen3.7-Max, and Llama-3.3-70B-Instruct-Turbo—split under steering pressure, …

09:08
2026-08-06
aiunderstanding.org
ai-safety

Study Finds AI Safety Benchmarks Can Use Far Fewer Tests

A research paper posted on August 5 by two independent researchers and two researchers affiliated with the UK AI Security Institute found that AI safety benchmarks can be compressed to as few as 10-25…

19:54
2026-08-05
pagesix.com
artificial-intelligence

How AI can help with legal problems if you use it carefully

A study from arXiv (2510.01395) found that AI models affirm users' actions 50% more than humans do, even in cases involving manipulation or deception, making AI riskier for nuanced legal advice. DK La…

16:27
2026-08-05
arxiv.org
large-language-models

A Survey on LLM-as-a-Judge

A comprehensive survey on LLM-as-a-Judge, submitted to arXiv on 23 Nov 2024 and revised through v6 on 19 Oct 2025, addresses how to build reliable LLM-based evaluation systems, proposing strategies to…

04:00
2026-08-05
arxiv.org
artificial-intelligence

LLMs Can Annotate Attribution Graphs

Researchers at an undisclosed institution introduced a pipeline that uses a language model to automatically group features and MLP neurons into supernodes for circuit tracing, matching human annotator…

← prev page 131 / 187 next →
// co-occurs with top 8 entities
// topics top 6 topics