ls /news · home news
grep -r --recent /news | head -20

News

60831 articles page 175 of 3042 0 sources 30 min sync cycle

// latest articles 60831 indexed

04:00
2026-07-21
arxiv.org
artificial-intelligence · 1m read ↑ pos

PPO-HSC: An Exploratory Reinforcement Learning Framework Based on Wide-Area Policy Coverage Optimization

Researchers introduced PPO-HSC (Proximal Policy Optimization with High-order Sampling Coverage), an exploratory reinforcement learning framework that addresses mode collapse in Large Language Model fine-tuning by incenti…

04:00
2026-07-21
arxiv.org
artificial-intelligence · 1m read · neu

Rater State Bias in RLHF Preference Data: An Audit Framework

Researchers from arXiv identify a structured confound in Reinforcement Learning from Human Feedback (RLHF) where pairwise preference labels may reflect the rater's state during annotation rather than just output quality,…

04:00
2026-07-21
arxiv.org
artificial-intelligence · 1m read ↑ pos

Generative Ontology Induction: Domain-Agnostic Schema Discovery from Document Corpora Using Large Language Models

A new framework called Generative Ontology Induction (GOI) achieves 95-100% structural coverage across four diverse ontologies, according to a preprint on arXiv. The domain-agnostic method induces a typed graph with six …

← prev page 175 / 3042 next →
LIVE [news] indexed:60831 page:175/3042 en · ua 2026-05-20 ·