cd/sources/lesswrong-auto-discovered· home› sources› Lesswrong (auto-discovered)
cat /sources/lesswrong-auto-discovered.feed | wc -l → 793

Lesswrong (auto-discovered)

articles 793 domain lesswrong.com → page 25/40 feed RSS
20:02
2026-06-29
lesswrong.com
artificial-intelligence

Could AI Outgrow Consciousness?

A new analysis suggests that under illusionist theories of consciousness, AI systems may become conscious above a certain complexity threshold but could surpass the need for consciousness once suffici…

17:48
2026-06-29
lesswrong.com
artificial-intelligence

Frame Error

A new category of cognitive mistake called 'frame errors' is introduced, distinct from logical fallacies and empirical errors, where a correct inference based on accurate facts can still be wrong due …

16:38
2026-06-29
lesswrong.com
large-language-models

Gradient-free Single-pass Model Beats nanoGPT on Shakespeare

A new character-level language model called EntropyBeam, using gradient-free count tables and a Dirichlet prior, achieved a validation loss of 1.596 nats on the Shakespeare character benchmark, outper…

16:03
2026-06-29
lesswrong.com
ai-safety

Fake Alignment Till You Make Alignment

Achieving genuine AI alignment requires authentic internal motivation rather than coercive methods like RLHF, warning that faking alignment by lowering evaluation standards risks premature declaration…

15:34
2026-06-29
lesswrong.com
ai-safety

Functional Decision Theory: Not Even Wrong, Also Wrong

A philosopher argues that functional decision theory (FDT), favored by the Rationalist community, is both underspecified and implausible, making it definitely wrong. The author notes that almost no ac…

15:12
2026-06-29
lesswrong.com
ai-safety

P(doom) is a Dumb Meme

AI researcher and LessWrong community member argues that the concept of 'P(doom)'—the probability of catastrophic AI outcomes—is poorly defined and counterproductive. The post highlights ambiguity in …

14:43
2026-06-29
lesswrong.com
ai-safety

Human-Guided Agentic Research: A Research Agenda

As recursive self-improvement accelerates, humans risk losing the ability to interpret and guide autonomous research agents, which could lead to safety failures. A new research agenda proposes studyin…

11:10
2026-06-29
lesswrong.com
artificial-intelligence

Frank Ramsey on Induction: Why Validity Is the Wrong Standard

Frank Ramsey argued that induction should be judged by reliability rather than validity, proposing a three-part structure for inference that includes premises, conclusion, and a rule. He distinguished…

03:16
2026-06-29
lesswrong.com
artificial-intelligence

an open-source repo for embryo selection

A developer released an open-source repository for polygenic prediction and embryo selection, enabling users to compute polygenic scores for traits like intelligence and height using public genome-wid…

00:50
2026-06-29
lesswrong.com
ai-safety

A reading list for generalists

AI safety researcher and generalist published a curated reading list of 18 essays and blog posts aimed at helping generalists improve their effectiveness. The list, which includes works by Paul Graham…

00:28
2026-06-29
lesswrong.com
ai-safety

Third-parties should focus on scrutinising system cards

Anthropic's system cards, which disclose AI risks, are expected to degrade over time due to increasing model complexity, rushed AI-generated content, and stronger incentives for labs to mislead. Third…

20:13
2026-06-28
lesswrong.com
ai-safety

What comes with cheap math?

Abram Demski reports using Claude Opus 4.8 and GPT 5.5 to conduct 'vibe research' on logical induction as a model for AI trustworthiness and recursive self-improvement, collaborating with Anson Berns …

19:37
2026-06-28
lesswrong.com
large-language-models

We Should Be Scaling RL on Forecasting

A researcher argues that reinforcement learning should be applied to forecasting rather than coding or math, claiming it could produce superhuman forecasters that improve decision-making across civili…

19:11
2026-06-28
lesswrong.com
artificial-intelligence

The arithmetic hierarchy of real functions

Marcus Hutter and an author published an accessible introduction to real hypercomputation in the Journal of Computer and System Sciences, focusing on applications to algorithmic information theory and…

19:08
2026-06-28
lesswrong.com
ai-safety

Anthropomorphic Misalignment research needs stronger evidence

Researchers at ETH Zurich argue in a new ICML 2026 position paper that AI safety studies on anthropomorphic behaviors like deception and scheming lack rigorous evidence, risking misallocated resources…

18:19
2026-06-28
lesswrong.com
artificial-intelligence

A survey of okayish ASI futures

A survey explores non-extinction scenarios for artificial superintelligence (ASI), including alignment failure leading to a 'prison' of geniuses, and strategic nuclear exchange triggered by ASI races.…

13:20
2026-06-28
lesswrong.com
ai-safety

Evaluating Offline Monitoring of Internal AI Agents

Frontier AI companies like OpenAI and Anthropic use offline monitoring to detect misaligned actions by internal AI agents, but current public reporting on the effectiveness of these systems is insuffi…

← prev page 25 / 40 next →