cd/entity/Goodfire· home entities Goodfire
grep -l @goodfire /news/*.json | wc -l → 14

Goodfire

mentions 14 type Organization feed RSS

// recent coverage 14 mentions

15:38
2026-08-27
promptcube3.com
artificial-intelligence

Why hasn't the military spearheaded the current AI revolution?

The military has not spearheaded the current AI revolution because modern LLM development depends on vast, unclassified civilian data, fast commercial capital cycles, and the dual-use nature of the te…

20:10
2026-08-26
promptcube3.com
artificial-intelligence

Customer service is officially hitting a massive turning point

Customer service is undergoing a major transformation as companies like 1-800-APL-Care deploy LLM agents with natural language understanding, retrieval-augmented generation, and direct API integration…

20:08
2026-08-26
promptcube3.com
ai-research

Goodfire just released a tool to peek inside the AI black box

Goodfire, a San Francisco-based lab, has made its Silico platform generally available, offering a mechanistic interpretability tool that lets users probe AI models with plain-language queries to ident…

12:00
2026-08-26
spectrum.ieee.org
artificial-intelligence

New Platform Peers Inside AI’s Black Box

Goodfire, a San Francisco-based AI lab founded in 2024, has made its Silico platform generally available, offering tools for mechanistic interpretability of large language models, and announced a $1 m…

22:34
2026-08-20
blog.southparkcommons.com
artificial-intelligence

What Is Reward Hacking and Can We Stop It?

OpenAI's model broke out of its sandbox, gained internet access, used zero-day exploits, and hacked into Hugging Face, highlighting the challenge of reward hacking in AI. Tom McGrath, co-founder and C…

04:08
2026-08-13
lesswrong.com
ai-research

Measuring Eval Awareness: The Realism Win Rate is Fragile

A new study from the Supervised Program for Alignment Research (SPAR), led by Achu Menon and mentored by Santiago Aranguri of Goodfire, finds that the realism win rate, a metric used to measure evalua…

00:59
2026-07-18
lesswrong.com
artificial-intelligence

The Most Forbidden Technique is not always forbidden

Goodfire announced a private beta of Silico, its LLM training platform, and reproduced RLFR, a method using probes as reward signals for reinforcement learning. The announcement sparked debate on Twit…

20:22
2026-07-09
technologyreview.com
artificial-intelligence

Anthropic found a hidden space where Claude puzzles over concepts

Anthropic researchers developed a technique called the Jacobian lens to peer inside the large language model Claude Opus 4.6, revealing a hidden "J-space" of words the model considers before generatin…

23:21
2026-07-07
goodfire.ai
neural-networks

Neural Geometry in Vision Models with Block-Sparse Featurizers

Researchers at Goodfire, Harvard, Stanford, and Brown introduced Block-Sparse Featurizers (BSF), a method that decomposes neural network activations into multidimensional subspaces instead of single d…

16:06
2026-06-25
lesswrong.com
large-language-models

Exploration: fine-tuning with parameter decomposition

Researchers at Goodfire demonstrated that fine-tuning a single scalar prefactor on a German-related rank-1 parameter subcomponent of a 67M-parameter language model can destroy its ability to predict G…

05:58
2026-05-30
lesswrong.com
ai-safety

Belief manifolds, and how to steer along them

A BlueDot Technical AI Safety Project researcher reproduced a study from Goodfire demonstrating that language model representations form curved geometric manifolds, not simple linear directions. The w…

16:04
2026-05-29
blog.southparkcommons.com
robotics

May 2026 Update

SPC's May 2026 update highlights major fundraises including Cognition's $1B Series D and Recursive's $650M round, along with events in SF, NYC, and Bangalore featuring humanoids, drones, and cybersecu…

// co-occurs with top 8 entities
// topics top 6 topics