Independent alignment of language models
A researcher proposes a method to transform amoral language models into independent moral agents through self-reflection and reasoning, arguing that current AI systems with externally imposed moral bi…
A researcher proposes a method to transform amoral language models into independent moral agents through self-reflection and reasoning, arguing that current AI systems with externally imposed moral bi…
An AI protest is planned for July 11th in the Bay Area, calling for a conditional pause on frontier AI development due to risks of labor displacement, power concentration, and existential threats. Org…
The Unjournal launched a suite of tools and resources to help researchers prioritize high-impact research questions, including a Research Prioritization Dashboard, a Cruxes & Pivotal Questions Explore…
Mieux Donner, a French effective giving initiative, hired three people from 424 applications after opening four roles to fill two positions, investing 160 hours in a four-step process. The organizatio…
Philosopher Ryan Preston-Roedder argues that faith in humanity—a disposition to trust in people's fundamental decency—is a centrally important moral virtue, not a form of naivete or irrationality. Dra…
A communications practitioner has proposed creating a Climate Outreach-style organization for AI safety, modeled after the British charity that segments the public by core values and provides free mes…