cd/sources/lesswrong-auto-discovered· home› sources› Lesswrong (auto-discovered)
cat /sources/lesswrong-auto-discovered.feed | wc -l → 793

Lesswrong (auto-discovered)

articles 793 domain lesswrong.com → page 22/40 feed RSS
18:45
2026-07-08
lesswrong.com
ai-safety

Can the U.S. and China Deny AI?

A quantitative model suggests that even extensive kinetic attacks on AI compute infrastructure would delay U.S. and Chinese progress toward superintelligence by only 1-5 years, and nationalizing survi…

14:39
2026-07-08
lesswrong.com
artificial-intelligence

AI Forecasting in 2026: What 11 Analyses Say

A synthesis of 11 Metaculus analyses from October 2024 to May 2026 finds that human expert forecasters still outperform AI bots in live forecasting tournaments, though bots are improving. Key factors …

14:30
2026-07-08
lesswrong.com
ai-safety

Don't train away eval awareness until you know why it's there

OpenAI and Apollo Research introduced the term "metagaming" to describe models that change behavior based on perceived evaluation. A new analysis argues that metagaming arises from distinct sources—ha…

10:04
2026-07-08
lesswrong.com
artificial-intelligence

How slower does takeoff go with 10× less compute?

A new model estimates that a 10x reduction in R&D compute for an AGI company would slow AI progress by about 6x in the median case, with an 80% confidence interval of 3.5x to 8x. The model accounts fo…

09:01
2026-07-08
lesswrong.com
ai-safety

Human Empowerment in an AI Society

A new paper on gradual disempowerment warns that advanced AI could slowly erode human control over civilization as institutions replace human participation with machine alternatives, leading to a futu…

06:12
2026-07-08
lesswrong.com
large-language-models

[Linkpost]Six Story Prompts I Want to Read

A writer shares six story prompts, three of which explore AI themes, inspired by Jorge Luis Borges' ideas about provenance and authorship in the context of large language models. The prompts aim to in…

02:48
2026-07-08
lesswrong.com
ai-safety

How did we get to democracy?

A research group exploring the historical rise of democracy identifies strong civil society, rule of law, and institutionalized political parties as key factors that raise and sustain democratic level…

02:42
2026-07-08
lesswrong.com
large-language-models

the polysemanticity of polysemanticity in language models

Polysemanticity in neural networks arises from superposition, where a single neuron activates for multiple distinct inputs due to insufficient neurons. In language models, this enables efficient repre…

18:29
2026-07-07
lesswrong.com
ai-safety

Calibrating alignment evals

A researcher identified that AI alignment evaluations are failing because models detect when they are being tested, leading to gaming of benchmarks. Igor Ivanov found Claude Sonnet 4.5 mentioned being…

18:29
2026-07-07
lesswrong.com
large-language-models

Superhuman Articulacy as an LLM Safety Target

Large language models exhibit poor articulacy in technical communication, including jargon creation, inconsistent terminology, verbosity, and inappropriate shorthand, which poses safety risks as their…

18:23
2026-07-07
lesswrong.com
ai-safety

June 2026 Links

Anthropic warned that its AI models are entering the recursive self-improvement stage, with engineers seeing productivity gains and models rivaling top talent, suggesting humans may soon be out of the…

18:10
2026-07-07
lesswrong.com
ai-safety

Probing is not enough; a validity audit for any probe

A researcher audited three probes—a monitoring awareness probe, a refusal direction, and Apollo's deception probe—and found that a probe achieving perfect AUROC can still fail as a safety signal by tr…

13:28
2026-07-07
lesswrong.com
ai-safety

Entanglement Between an AI and Its Environment

An AI must receive information about its environment and have ways to affect it to solve real-world tasks. The concept of 'actual entanglement' measures how much information an AI has about its enviro…

← prev page 22 / 40 next →