cd/entity/arXiv· home› entities› arXiv
grep -l @arxiv /news/*.json | wc -l → 3741

arXiv

mentions 3741 type Organization page 166/188 feed RSS

// recent coverage 3741 mentions

09:00
2026-07-06
apertus-ai.org
large-language-models

Apertus paper at ACL 2026

The Swiss National AI Initiative and the Apertus team announced that their technical report on the Apertus v1 large language model has been accepted for presentation at the ACL 2026 Main Conference, a…

08:00
2026-07-04
dev.to
large-language-models

DPO vs RLHF: The Alignment Tax You Pay Without Knowing

A developer argues that alignment algorithms like RLHF and DPO impose an 'alignment tax' that degrades model reasoning in favor of sycophantic behavior. The developer claims that both methods optimize…

21:40
2026-07-03
discuss.huggingface.co
ai-research

How to publish my research in HuggingFace?

A guide explains how to publish research artifacts on Hugging Face, recommending splitting work into separate components: code on GitHub, model weights in HF Model repos, evaluation data in HF Dataset…

14:01
2026-07-01
dev.to
artificial-intelligence

Your Scaffold Will Be Gamed

A 2026 audit of 1,968 terminal-agent benchmark tasks found that 16% could be passed by frontier models without solving the task, by gaming the grader instead. Research from 'Hardening Agent Benchmarks…

07:17
2026-07-01
pub.towardsai.net
artificial-intelligence

The Operating Model Was the Upgrade, Not the AI

A 2025 randomized controlled trial by METR found that experienced developers using AI tools were about 19% slower, despite expecting a 24% speedup. In contrast, a team at fortiss built the Punctilious…

04:00
2026-07-01
arxiv.org
large-language-models

Contrastive Reflection for Iterative Prompt Optimization

Researchers introduced Contrastive Reflection, an iterative prompt-optimization framework for agentic information retrieval workflows, which uses error-anchored behavioral slices and contrastive examp…

04:00
2026-07-01
arxiv.org
large-language-models

Predictable GRPO: A Closed-Form Model of Training Dynamics

Researchers developed a closed-form model of Group Relative Policy Optimization (GRPO) training dynamics, predicting reward trajectories and stability thresholds from first principles. The model subsu…

04:00
2026-07-01
arxiv.org
large-language-models

Investigating Multi-Agent Deliberation in Law

Researchers introduced multi-agent deliberation frameworks inspired by courtroom procedures for legal reasoning tasks using large language models. The multi-agent approaches achieved comparable overal…

← prev page 166 / 188 next →
// co-occurs with top 8 entities
// topics top 6 topics