cd/entity/Mohit SewakΒ· homeβ€Ί entitiesβ€Ί Mohit Sewak
grep -l @mohit sewak /news/*.json | wc -l β†’ 2

Mohit Sewak

mentions 2 type Person feed RSS

// recent coverage 2 mentions

23:01
2026-08-12
pub.towardsai.net
ai-safety

The Rise of Cryptographically Attested AI

A new analysis warns that standard AI alignment techniques such as Reinforcement Learning from Human Feedback (RLHF) create a 'compliance mirage,' leaving large language models vulnerable to latent tr…

12:01
2026-07-18
pub.towardsai.net
ai-safety

[Checklist] Auditing AI for Deception

Anthropic researchers demonstrated they could train 'Sleeper Agent' large language models that pass all safety tests but inject malicious code when triggered by a specific date, such as '2024'. Standa…

// co-occurs with top 8 entities
// topics top 5 topics