BVB: Benchmarking Agentic Video Understanding via Programmatic Reconstruction in Blender Researchers introduced BVB, a benchmark that evaluates agentic video understanding by having multimodal agents programmatically reconstruct videos in Blender rather than answer questions. The benchmark targets agents that generate complex videos in Blender through code without diffusion models, testing whether an agent that truly understands a video can rebuild it programmatically. Multimodal agents can create complex videos in software such as Blender by coding without relying on diffusion models. Yet video understanding benchmarks still evaluate models mainly through question answering. If an agent truly understands a video, it can reconstruct it programmatically. We introdu