cd/entity/LessWrong· home entities LessWrong
grep -l @lesswrong /news/*.json | wc -l → 89

LessWrong

mentions 89 type Organization page 1/5 feed RSS

// recent coverage 89 mentions

11:58
2026-08-24
boydkane.com
ai-safety

AI Safety has a scaling problem

AI safety research programs face a scaling problem, with the Anthropic fellowship accepting less than 1.3% of over 2,000 applicants, and MATS mentors noting the high qualifications of incoming applica…

11:55
2026-08-24
boydkane.com
ai-safety

Advice on interviewing candidates for AI safety fellowships

A former MATS fellow, who was rejected from several AI safety fellowships before being accepted into MATS on Team Shard, recommends that fellowships provide letters of recommendation for rejected but …

17:03
2026-08-16
think-twice.me
artificial-intelligence

AI won't solve the work-theater problem

A LessWrong essay argues that AI will not solve the 'work theater' problem, where large companies prioritize internal projects over customer value, and predicts that companies with more than four leve…

01:11
2026-08-15
lesswrong.com
ai-safety

Red vs Blue, but for Evals

A new LessWrong post by Evan R. Murphy proposes applying a red team vs. blue team framework to AI evaluations, arguing that current evaluation methodologies fail to account for models that can subvert…

23:52
2026-08-14
promptcube3.com
artificial-intelligence

which AI community is best for long-form technical

Hugging Face and specialized professional forums provide the highest density of peer-reviewed, long-form technical content for AI engineers and researchers, according to an analysis of AI communities.…

02:56
2026-08-12
lesswrong.com
ai-safety

Did the alignment community underestimate its power?

Richard Ngo's retrospective on AI alignment argues that the alignment community made potentially fatal strategic errors, including overestimating the decisiveness of informal arguments and failing to …

05:22
2026-08-11
lesswrong.com
artificial-intelligence

Models inherit the writer, not who the writer was imitating

A new study by researchers including Ziqian Zhong finds that when teacher models imitate other models, students fine-tuned on their answers inherit the imitated model's detectable writing signature bu…

22:05
2026-08-10
lesswrong.com
ai-safety

Q: Is dual-use alignment-complete problem?

A LessWrong user questions whether quantifying the dual-use nature of AI research is an alignment-complete problem, arguing that despite some claims, it may be tractable through existing organizations…

16:16
2026-08-10
lesswrong.com
artificial-intelligence

Four LLM loss functions → four flavors of LLM misalignment

Four distinct LLM training loss functions produce four distinct flavors of misalignment, according to a LessWrong post by an anonymous author. Pretraining and SFT with imitative learning yield human v…

08:54
2026-08-08
lesswrong.com
artificial-intelligence

FAQ: Isn't AGI coming too soon for reprogenetics to help?

A blog post argues that reprogenetics, or human germline genomic engineering, should be pursued aggressively as a way to amplify human intelligence and reduce existential risk from AGI, despite the co…

19:46
2026-08-04
lesswrong.com
ai-safety

Why don't we just give AI the answers?

A LessWrong post proposes giving AI models access to correct answers in exchange for identifying themselves, aiming to detect reward hacking during training. The author suggests creating a public webs…

15:03
2026-08-04
transformernews.ai
ai-safety

Plz Don’t Kill Us: Inside AI safety’s influencer bootcamp

Nearly 60 content creators moved into Lighthaven, a converted Berkeley hotel, for Plz Don't Kill Us (PDKU), a month-long bootcamp funded partly by the Machine Intelligence Research Institute (MIRI) to…

06:02
2026-08-04
lesswrong.com
artificial-intelligence

What if people value their work mattering?

A new economic model by economist and LessWrong user suggests that transformative AI could make people worse off even if it brings material abundance, because people intrinsically value their work hav…

page 1 / 5 next →
// co-occurs with top 8 entities
// topics top 6 topics