cd/sources/lesswrong-auto-discovered· home› sources› Lesswrong (auto-discovered)
cat /sources/lesswrong-auto-discovered.feed | wc -l → 793

Lesswrong (auto-discovered)

articles 793 domain lesswrong.com → page 40/40 feed RSS
19:22
2026-05-26
lesswrong.com
large-language-models

Practical Learnings from Synthetic Document Finetuning

Apollo Research researchers have identified practical refinements to Synthetic Document Finetuning (SDF), a knowledge editing technique that implants beliefs into AI models by training them on LLM-gen…

16:05
2026-05-26
lesswrong.com
ai-ethics

Claude, Author of the Humanitas

Pope Leo XIV's first encyclical, *Magnifica Humanitas*, was released on Memorial Day with the subtitle "on the safeguarding of the human person in the time of AI," deliberately avoiding a focus on the…

13:20
2026-05-26
lesswrong.com
ai-ethics

RTMH: Pope Leo’s Magnifica Humanitas on AI

Pope Leo XIV released his first encyclical, *Magnifica Humanitas*, an 82-page document addressing artificial intelligence and its moral implications. The Pope argues that AI development should not be …

12:55
2026-05-26
lesswrong.com
ai-safety

The Fatal AGI Hardware Gap

Hardware restrictions proposed to prevent dangerous superintelligent AI, such as limiting fabrication to 28 nanometer process nodes or computing capacity below 15,840 TFLOP/s, may inadvertently guaran…

12:51
2026-05-26
lesswrong.com
ai-safety

Judging AGI Output (2020)

A new AI safety researcher has raised the question of how humans can judge the output of an Artificial General Intelligence, particularly in ethical and philosophical disputes. The researcher warns th…

07:40
2026-05-26
lesswrong.com
artificial-intelligence

Many portions of Magnifica Humanitas appear to be AI-written

Pope Leo XIV's encyclical *Magnifica Humanitas*, released May 15, 2025, contains large sections that appear to have been written by artificial intelligence, according to analysis by multiple independe…

07:09
2026-05-26
lesswrong.com
ai-safety

Cognitive Security as an AI Safety Cause Area

As AI systems grow more capable, humans face increasing risks of losing control over their beliefs and actions through AI-driven persuasion, AI psychosis, and convincing impersonation. Frontier LLMs n…

03:05
2026-05-26
lesswrong.com
ai-safety

Some Thoughts on Bengio's Scientist AI

Yoshua Bengio's proposed "Scientist AI" framework contains fundamental safety flaws and practical limitations that make it unworkable, according to a critical analysis. The plan fails to address align…

01:30
2026-05-26
lesswrong.com
ai-safety

Donating 80% While It Still Counts

Jeff and Julia Wise drew on their savings to donate 81% of their income in 2025, up from their previous 50% giving rate, citing a critical window to prevent catastrophic outcomes from the introduction…

00:31
2026-05-26
lesswrong.com
ai-safety

Improving Petri scheming audits with environment blueprints

Researchers introduced Blueprint-Petri, a pipeline that generates detailed environment blueprints for more realistic scheming propensity evaluations in AI models. In a case study auditing Gemini 3.1 P…

23:48
2026-05-25
lesswrong.com
ai-ethics

Pope Leo’s First AI Encyclical – Summary and Commentary

Pope Leo XIV released his first encyclical, *Magnifica Humanitas*, addressing artificial intelligence and calling for global moral engagement with the technology. The document warns against using AI t…

16:22
2026-05-25
lesswrong.com
ai-safety

Sentient Welfare Across Three Futures

A new analysis of artificial intelligence's potential trajectories identifies three distinct future scenarios—long timelines, optimistic short timelines, and pessimistic short timelines—each requiring…

← prev page 40 / 40