cd/entity/arXiv· home› entities› arXiv
grep -l @arxiv /news/*.json | wc -l → 3625

arXiv

mentions 3625 type Organization page 54/182 feed RSS

// recent coverage 3625 mentions

00:15
2026-09-15
dev.to
ai-agents

Why you shouldn't let the model review its own AI code

A new arXiv paper by Krentsel, Agarwal, Cemri, Zaharia and Stoica, "Reality Is the Final Verifier: On Two Key Gaps in Agentic Software Engineering," argues that having the same model review the code i…

21:25
2026-09-14
discuss.huggingface.co
ai-agents

A thought about AI models working together as a pipeline

Research papers including FilmAgent, VideoGen-of-Thought, CineAGI, StoryAgent, and AniME are decomposing video production into specialized agents coordinated by an orchestrator, with CineAGI reporting…

07:32
2026-09-14
snipvote.com
ai-agents

Native harnesses don't always solve more coding tasks

A study of 256 private coding tasks found no reliable overall performance advantage for vendor-native harnesses over neutral third-party harnesses on the same models, with Opus 4.8 scoring 48.8% versu…

07:32
2026-09-14
snipvote.com
ai-agents

Occamy-1.0: Open Pareto-frontier 35B Intelligence for Co-work

Occamy-1.0, an open 35B co-work agent model trained from Qwen3.6-35B-A3B, lands at the low-cost knee of the cost-performance Pareto frontier across four representative co-work benchmarks, according to…

04:28
2026-09-14
leimao.github.io
machine-learning

Residual-Quantized Variational Autoencoder

A blog post explains the Residual-Quantized Variational Autoencoder (RQ-VAE), an architecture that extends the Vector Quantized Variational Autoencoder (VQ-VAE) introduced in arXiv paper 1711.00937 to…

04:00
2026-09-14
arxiv.org
ai-agents

Look Before You Leap: Pre-Action Verification for LLM Agents

A new arXiv paper (2609.11957v1) proposes pre-action verification as an underused form of LLM agent oversight, testing it across shell commands and code edits. For shell commands, a static verifier ov…

04:00
2026-09-14
arxiv.org
machine-learning

Efficient AI Model Deployment Using Quantization Analysis Tool

A new arXiv paper (2609.11954v1) presents Quantization Analysis Tool, a system built on the ONNX framework that streamlines quantization workflows for deploying deep learning models on resource-constr…

← prev page 54 / 182 next →
// co-occurs with top 8 entities
// topics top 6 topics