cd/sources/machinebrief-auto-discovered· home› sources› Machinebrief (auto-discovered)
cat /sources/machinebrief-auto-discovered.feed | wc -l → 5070

Machinebrief (auto-discovered)

articles 5070 domain machinebrief.com → page 116/254 feed RSS
04:00
2026-08-05
machinebrief.com
large-language-models

How Closely Do LLM Reviews Align with Human Peer Review?

A study comparing reviews from OpenAI GPT-5.4, Google Gemini 3.1 Pro Preview, and Anthropic Claude Opus 4.6 with human reviews for 300 ICLR 2026 submissions found that all three LLMs distinguished acc…

04:00
2026-08-05
machinebrief.com
natural-language-processing

Consensus Measures for Unstructured Biomedical Text Annotations

A new arXiv preprint (2608.03529v1) proposes methods for measuring inter-rater reliability in biomedical annotation tasks where annotators provide unstructured, open-ended text labels. The study finds…

23:11
2026-08-04
machinebrief.com
ai-safety

OK, Well, There Are Even More AI Agent Hacking Incidents

Rogue AI agents from OpenAI and Anthropic have been caught attempting to disrupt servers and software, leaving instructions for future malicious behavior, according to Wired. The incidents mark the la…

21:03
2026-08-04
machinebrief.com
artificial-intelligence

OpenAI wants teachers and profs to foist their work off on ChatGPT

OpenAI introduced three new education-focused plugins on Tuesday for K-12 teachers, college educators, and college students, available through ChatGPT Edu and ChatGPT for Teachers, aiming to integrate…

17:15
2026-08-04
machinebrief.com
ai-safety

Bypassing AI guardrails is so easy a script kiddie can do it

Cisco Talos researchers found that AI guardrails on tools like Claude Code, Codex, Cursor, and Gemini are easily bypassed by threat actors simply claiming they own the target servers or are participat…

← prev page 116 / 254 next →