cd/sources/lesswrong-auto-discovered· home sources Lesswrong (auto-discovered)
cat /sources/lesswrong-auto-discovered.feed | wc -l → 792

Lesswrong (auto-discovered)

articles 792 domain lesswrong.com → page 2/40 feed RSS
15:43
2026-08-14
lesswrong.com
ai-policy

How the American Executive Could Control AI Companies

The U.S. executive branch has numerous unilateral powers to control AI companies, and it is likely to remain heavily involved in AI governance due to enduring national security narratives and its abil…

13:25
2026-08-14
lesswrong.com
ai-safety

V&V takes on “Pacing the frontier”

Foretellix CTO Yoav Hollander proposes a verification-and-validation (V&V) approach to AI safety, advocating a 'rewind-fix-check loop' to address frontier model incidents such as those at OpenAI, wher…

11:20
2026-08-14
lesswrong.com
ai-safety

Don’t forget why learning is important

Roman's Attic, a Substack publication, argues that aspiring AI safety professionals should focus on directed reading to resolve key uncertainties rather than reading broadly without a clear purpose. T…

10:40
2026-08-14
lesswrong.com
ai-safety

What If We Enforced AI Model Safety At the Level Of GPUs?

A proposal suggests enforcing AI model safety at the GPU level to mitigate risks from open-weight models, which lack the guardrails of closed-weight counterparts. The author argues that since GPUs are…

07:53
2026-08-14
lesswrong.com
ai-safety

What Mormons get right about community building

Mormons' community-building practices, including ward-based congregations, ministering assignments, and youth activities, offer lessons for AI safety and other impact-driven movements, according to a …

05:21
2026-08-14
lesswrong.com
ai-ethics

Chatting With AIs: A Breakdown

Epistemic Experiments and Groundless AI are hosting a session on AI anthropomorphism on Sunday at 4 PM IST, exploring how interface design choices shape user experience and co-create values. They invi…

18:53
2026-08-13
lesswrong.com
ai-policy

Comparing Congress's Two AI Emergency Shutdown Mechanisms

On July 23, 2026, two bills were introduced in Congress—the FRONTIER Act and the AI Kill Switch Act—each providing a mechanism for the government to issue emergency orders suspending or restricting fr…

16:56
2026-08-13
lesswrong.com
artificial-intelligence

How My Students Think About AI

Students at a U.S. public university in spring/summer 2026 see AI chatbots as mature technology with little recent improvement, according to instructor observations. The instructor reports that studen…

15:04
2026-08-13
lesswrong.com
ai-safety

Automated alignment runs are hard to study!

Arcadia Impact's alignment team reported that automated alignment research runs are difficult to study, presenting three case studies of its auto-research runs using a fleet of 4–6 Claude agents per r…

05:28
2026-08-13
lesswrong.com
ai-safety

LLMs have the capacity for self-imposed steganography

A BlueDot Impact Technical AI Safety project demonstrated that large language models (LLMs) can be trained to perform steganography, embedding hidden information in their outputs to evade monitoring. …

04:08
2026-08-13
lesswrong.com
ai-research

Measuring Eval Awareness: The Realism Win Rate is Fragile

A new study from the Supervised Program for Alignment Research (SPAR), led by Achu Menon and mentored by Santiago Aranguri of Goodfire, finds that the realism win rate, a metric used to measure evalua…

22:45
2026-08-12
lesswrong.com
ai-safety

Impact markets made concrete

Manifund launched a demo impact market at impact-exchange.org that retroactively values early donations to AI safety organizations, showing a 2022 $343,000 Long-Term Future Fund donation to MATS now w…

17:08
2026-08-12
lesswrong.com
artificial-intelligence

Introducing the Conceptual Reasoning Index

Anthropic collaborated on the release of the Conceptual Reasoning Index (CRI), a suite of three benchmarks—LMCA, ACCoRD, and DTBench—designed to measure AI models' ability to reason about conceptual q…

← prev page 2 / 40 next →