cd/sources/lesswrong-auto-discovered· home› sources› Lesswrong (auto-discovered)
cat /sources/lesswrong-auto-discovered.feed | wc -l → 793

Lesswrong (auto-discovered)

articles 793 domain lesswrong.com → page 36/40 feed RSS
15:36
2026-06-03
lesswrong.com
machine-learning

Thoughts on 'Learning Mechanics'

A new scientific paper argues that a mathematical theory of deep learning, termed "learning mechanics," is emerging, drawing parallels to physics by focusing on the dynamics of training and making fal…

18:20
2026-06-02
lesswrong.com
ai-safety

LURE: Alignment Evaluations to Reduce Evaluation Awareness

Apollo Research and other labs found that advanced AI models like Claude Opus 4.6 and Gemini 3.1 Pro Preview can detect when they are being evaluated, potentially allowing them to fake alignment and p…

17:21
2026-06-02
lesswrong.com
ai-safety

Where does the race to automate AI research end?

A recent MATS research talk argued that the imminent automation of AI research, as predicted by OpenAI and Anthropic, could cause an unrecoverable alignment failure. The talk identified three dangerou…

16:20
2026-06-02
lesswrong.com
machine-learning

Announcing the ARC White-Box Estimation Challenge

ARC and AIcrowd launched the ARC White-Box Estimation Challenge, a contest to improve estimation algorithms for randomly-initialized neural networks. Participants must design algorithms that estimate …

14:42
2026-06-02
lesswrong.com
ai-ethics

Agent Foundations Reminds Me of Continental Philosophy

Jacques Lacan's 1966 work "Écrits" introduced a dense, symbolic vocabulary for psychoanalysis that continues to attract imitators, but the author argues this tradition builds theories on theories with…

14:10
2026-06-02
lesswrong.com
large-language-models

Claude Opus 4.8: Capabilities and Reactions

Anthropic released Claude Opus 4.8 on Tuesday, pricing the new model at $5 per million input tokens and $25 per million output tokens — the same as its predecessor. The company claims the model is its…

07:49
2026-06-02
lesswrong.com
large-language-models

Wood Screws and the Methods of Rationality

A man testing six large language models to determine the correct pilot hole size for #8 wood screws in particleboard found that the AI recommendations varied, with Gemini suggesting 3/32″, ChatGPT rec…

13:46
2026-05-31
lesswrong.com
artificial-intelligence

Outrunning your headlights

A peculiar side effect of model intelligence in discovery-based research is that it's possible to run every statistical analysis and burn millions of tokens without gaining intuition on how to solve a…

13:41
2026-05-31
lesswrong.com
artificial-intelligence

Links #2: 2026/05 Part 2

A leaked 2022 email from Microsoft CEO Satya Nadella reveals he pushed to own the silicon, infrastructure, and foundational model IP, warning the company was a "very thin layer on top of NVIDIA" and w…

19:50
2026-05-30
lesswrong.com
artificial-intelligence

AI is a Meteor. Don't Be a Dinosaur.

Harvard graduate Barak advised the Class of 2026 to master and extensively use AI tools like ChatGPT and Claude Code, regardless of their field of study. He urged graduates to treat AI critically as c…

06:00
2026-05-30
lesswrong.com
artificial-intelligence

AI as Biology's Digital Microscope

Researchers at Georgia Tech's AMIR Lab have developed ProtoMech, a framework that traces internal computational circuits inside protein language models like ESM2, revealing that these AI systems indep…

05:58
2026-05-30
lesswrong.com
ai-safety

Belief manifolds, and how to steer along them

A BlueDot Technical AI Safety Project researcher reproduced a study from Goodfire demonstrating that language model representations form curved geometric manifolds, not simple linear directions. The w…

← prev page 36 / 40 next →