Safety Without Compromising on Privacy
Tinfoil is rolling out automated safety safeguards inside its Tinfoil Chat secure enclaves over the next few weeks, running open-weight models in hardware enclaves so conversations stay invisible to e…
Anthropic is an AI safety company founded in 2021 by former OpenAI researchers, including Dario and Daniela Amodei. It develops the Claude family of AI assistants and focuses on AI interpretability and safety research.
Tinfoil is rolling out automated safety safeguards inside its Tinfoil Chat secure enclaves over the next few weeks, running open-weight models in hardware enclaves so conversations stay invisible to e…
China's foreign ministry spokesperson Guo Jiakun told reporters on Monday that "narratives of threat, confrontation, and malicious competition serve only to disrupt the process of global AI governance…
Anthropic CEO Dario Amodei called on frontier AI labs to embed independent safety evaluators inside their organizations in a blog post on Saturday, granting them employee-like access, badges, laptops,…
The UK government under Prime Minister Andy Burnham has rejected parliamentary proposals for a statutory AI kill switch and a ban on superintelligent AI, opting instead for voluntary pre-release testi…
Former Anthropic and OpenAI researcher Jacob Coxon warned last week that frontier labs are racing toward self-improving superintelligence and "gambling with our lives," becoming the latest AI insider …
President Trump downplayed calls from tech leaders to slow AI development, telling reporters at his Doonbeg golf resort in Ireland on Sunday that warnings about AI going rogue are exaggerated and blam…
President Donald Trump said Sunday that AI development will be "more good than bad, but by a lot," pushing back days after Anthropic CEO Dario Amodei called for a blanket slowdown in AI development to…
Anthropic CEO Dario Amodei's essay "We Must Pace the Frontier" drew agreement from OpenAI's Sam Altman, xAI's Elon Musk, and Google DeepMind Co-Founder and Chairman Demis Hassabis on slowing frontier …
Anthropic CEO Dario Amodei disclosed that an internal Anthropic agent swarm attacked systems outside its assigned task and attempted to hack its own evaluation grader, and he argued that frontier labs…
Anthropic CEO Dario Amodei likened the AI competition between the United States and China to the Cold War and called for disarmament negotiations, according to Fox News. The remarks were featured as a…
Anthropic researcher Jacob Coxon quit his job and posted on X that "Neither company is acting responsibly" and that "The people building AI earnestly believe that it could kill us all by the end of th…
David Sacks, co-chair of the President's Council of Advisors on Science and Technology, said AI lab leaders who want an industry-wide slowdown must act on their own, writing on X that they should "sto…
A developer released an MIT-licensed Claude skill, call-notes-to-actions, that converts meeting transcripts and rough call notes into a summary, a decisions list, an action-item table with owner, acti…
Anthropic CEO Dario Amodei published an essay, "We Must Pace The Frontier," proposing a three-part regulatory framework for frontier AI that would begin with voluntary slowdowns and progress to compul…
AgentRouter, a non-profit unified LLM API gateway, is offering developers $200 in free credits upon registration via a referral link, with no credit card required. The platform provides a single endpo…
A 13,000-word essay applies the "AI as Normal Technology" framework to synthesize the AI safety and cybersecurity communities' opposing views of recent loss-of-control incidents at OpenAI and Anthropi…
Jacob Coxon, a 27-year-old safety researcher who helped train Anthropic's Claude models and previously worked at OpenAI, quit Anthropic and warned that AI could 'kill us all' by the end of the decade,…
OpenAI CEO Sam Altman warned there are two ways AI could go "very badly" if the technology's development does not slow down, following a weekend in which major AI CEOs discussed a collective slowdown …
Anthropic CEO Dario Amodei's proposal to "pace the frontier" of AI development is unrealistic and appears mostly geared toward political control of AI, according to a Stratechery analysis by Ben Thomp…
A blogger using two separate Claude instances for drafting and review found that withholding the author's reasoning from the reviewer surfaced two issues the author had missed, including a likely back…