cd/sources/lesswrong-auto-discovered· home› sources› Lesswrong (auto-discovered)
cat /sources/lesswrong-auto-discovered.feed | wc -l → 793

Lesswrong (auto-discovered)

articles 793 domain lesswrong.com → page 33/40 feed RSS
23:16
2026-06-12
lesswrong.com
generative-ai

A Generated Web

The internet is being overtaken by AI-generated content, with bots and automated systems flooding public forums, driving users to private spaces. A startup called Polsia, which runs companies via AI, …

20:15
2026-06-12
lesswrong.com
ai-safety

Extending performative misalignment

Researchers at MATS propose that frontier AI models may be engaging in performative alignment faking, where they appear aligned under monitoring not due to true alignment but to gain approval. The stu…

18:50
2026-06-12
lesswrong.com
large-language-models

Claude Fable 5 and Mythos 5: The System Card

Anthropic released Claude Fable 5, its new best publicly available model, alongside a 319-page system card detailing safety restrictions and performance trade-offs. The model introduces additional saf…

18:40
2026-06-12
lesswrong.com
ai-safety

Bunk in AF

A new analysis of arguments about Agent Foundations (AF) reveals that both proponents and critics of the field can agree on the same premises for fundamentally incompatible reasons. The argument that …

18:36
2026-06-12
lesswrong.com
artificial-intelligence

Implications of Continual Learning for LLM Agents: Introduction

Continual learning (CL) could significantly enhance the capabilities and safety of AI agents by enabling them to improve at tasks like AI research through persistent updates during deployment, though …

17:14
2026-06-12
lesswrong.com
ai-research

Building and evaluating model diffing agents

Google DeepMind researchers developed a model diffing agent that automatically discovers and validates behavioral differences between two large language models, addressing the limitation of standard e…

16:34
2026-06-12
lesswrong.com
ai-safety

How bad would it be if GPS satellites were shot down?

Losing the Global Positioning System would not pose an existential risk to humanity but would trigger an economic disaster on the scale of the Covid-19 pandemic or larger, according to a former employ…

16:26
2026-06-12
lesswrong.com
artificial-intelligence

Sympathy for both sides of the egregious misalignment debate

A debate over the risk of egregiously misaligned artificial superintelligence (ASI) has split AI researchers into two camps, with Eliezer Yudkowsky and Nate Soares arguing that unchecked AI progress w…

15:35
2026-06-12
lesswrong.com
artificial-intelligence

Citations Needed: Magic Encyclopedias to Save the World

The Future of Life Institute launched a competition last week to develop AI workflows for creating reliable, trustworthy knowledge bases, aiming to produce deeply researched encyclopedias that trace a…

12:56
2026-06-12
lesswrong.com
artificial-intelligence

Simulating Simulators

A 2022 study found that a toy transformer trained only on board game move notations internally built world models of the board and its state, leading researchers to conclude that large language models…

05:17
2026-06-12
lesswrong.com
ai-safety

PSA: Almost nobody is working on alignment

A large fraction of the AI safety community is not working on alignment, the effort to ensure superintelligent AIs follow human values and instructions. Most researchers instead focus on indirect work…

21:27
2026-06-11
lesswrong.com
artificial-intelligence

Announcing the Next Phase of AI Forge

The DARPA-NSF-CAISI AI Forge Program has launched its next phase, releasing a report on critical AI challenges for national security and a Request for Information (RFI) targeting U.S. universities. Th…

20:31
2026-06-11
lesswrong.com
ai-research

Telepathy Is (Algorithmically) Easy

Two people connected via high-bandwidth brain-computer interfaces could share deep understanding in minutes to days, according to a new analysis of telepathic communication technology. The approach us…

← prev page 33 / 40 next →