04:00
2026-09-11
machinebrief.com
ai-safety
Understanding In-Context Multimodal Jailbreaks via Posterior Reweighting
A new arXiv paper (2609.10613v1) proposes a posterior reweighting framework that models safety-aligned multimodal large language models (MLLMs) as implicitly operating over competing behavioral modes,…