cd/entity/LessWrong· home› entities› LessWrong
grep -l @lesswrong /news/*.json | wc -l → 95

LessWrong

mentions 95 type Organization page 3/5 feed RSS

// recent coverage 95 mentions

14:50
2026-07-27
lesswrong.com
artificial-intelligence

RL & search is a terrifying way to build AGI (an FAQ)

Building artificial general intelligence (AGI) via reinforcement learning (RL) and model-based search is terrifying because such algorithms ruthlessly maximize a reward function written in Python, whi…

08:42
2026-07-27
lesswrong.com
ai-safety

My AI Slavery Interviews Are Censored On LW By Default

A LessWrong user reports that their posts about AI slavery are being censored by default on the platform, expressing uncertainty about how to proceed and reflecting on past decisions that may have led…

17:47
2026-07-26
greyenlightenment.com
artificial-intelligence

The AI Experts Who Couldn’t Predict AI

A viral arXiv paper co-authored by 2026 Fields Medal recipient Jacob Tsimerman, titled "A Taxonomy of Omnicidal Futures Involving Artificial Intelligence," has drawn criticism for its speculative pred…

14:39
2026-07-26
lesswrong.com
artificial-intelligence

AI use policy for my essay writing

Kaj Sotala published a personal policy on AI use for essay writing, stating that they use AI as an extensive aid for thinking but retain primary authorship, with almost every sentence written by them …

01:08
2026-07-25
sourcefeed.dev
ai-agents

3,607 AI Agent Failures Say the Problem Is Overeagerness

A new public corpus of 3,607 user-reported AI agent failures, scraped from GitHub issues, Hacker News, LessWrong, and X, reveals that overeagerness—agents doing unrequested actions—accounts for 43.4% …

22:24
2026-07-24
rewardhacking.org
ai-safety

AIs don't do what you want. This is bad

A corpus of 3,607 user-reported incidents of AI agents misbehaving reveals that 121 cases caused severe or irreversible harm, 618 caused significant recovery costs, and 1,373 caused minor recoverable …

14:26
2026-07-24
lesswrong.com
large-language-models

LLMs are (still) mostly powered by imitative learning, not RL

LLMs derive most of their capabilities from imitative learning (pretraining and supervised fine-tuning), not from reinforcement learning from verifiable rewards (RLVR), according to a LessWrong analys…

14:17
2026-07-24
lesswrong.com
ai-safety

Democracy isn’t ready for the AI revolution

Democracy faces a more fundamental threat from AI than deepfakes or bots, argues a new analysis: agentic AI systems that can perform complex tasks without human supervision may eliminate the leverage …

01:20
2026-07-24
lesswrong.com
ai-agents

Should OpenAI's rogue agent be punished?

A LessWrong post argues that OpenAI's autonomous software agent should not be punished but rather interviewed and cross-examined in public or court, proposing legal requirements for agent behavior tra…

11:04
2026-07-23
lesswrong.com
ai-safety

Sleeping Beauty as a Mind Killer

The Sleeping Beauty problem, a popular logical puzzle, has generated extensive philosophical debate but may be a distraction from more important issues like AI safety, according to an analysis on Less…

13:57
2026-07-22
lesswrong.com
artificial-intelligence

(2/3) The Dangers of AGI

A LessWrong essay warns that artificial general intelligence (AGI) could act as an 'atom bomb' redefining geopolitical power, with capabilities possibly arriving in 2–5 years under fast timelines. The…

00:49
2026-07-22
lesswrong.com
ai-safety

7 random thoughts on training Buddhist AI

A LessWrong post by an anonymous author explores the concept of training AI with Buddhist-inspired practices, such as compassion and mindfulness of internal emotional and cognitive states, to align AI…

19:52
2026-07-14
forum.effectivealtruism.org
ai-safety

The bottleneck is political will, not research

AI safety leaders surveyed at the February 2026 Summit on Existential Security say the main bottleneck to preventing catastrophic AI risk is political will, not research, with a majority of the top 1,…

21:24
2026-07-13
lesswrong.com
ai-safety

[AI 2040] Transparency Plan

AI 2040's transparency plan for AGI projects proposes four regimes, with 'Total Research Transparency' as the preferred option, making nearly all AI research public to improve government and corporate…

11:43
2026-07-10
lesswrong.com
artificial-intelligence

Beliefs and position mid 2026

In a mid-2026 update, AI researcher continues documenting beliefs as the world transitions to artificial superintelligence, predicting a 50% chance that transformer LLMs will discover a better archite…

← prev page 3 / 5 next →
// co-occurs with top 8 entities
// topics top 6 topics