PyTorch 2.13: FlexAttention on Apple Silicon, 4x Memory Savings, Upgrade Guide
PyTorch 2.13 shipped July 8 with FlexAttention gaining native Metal support on Apple Silicon, delivering up to 12x faster performance than SDPA on sparse patterns, and a new fused loss function that c…