Understanding PyTorch’s Test Infrastructure
PyTorch's test infrastructure dynamically generates test names across devices and dtypes, causing CI failures to show names like TestLinalgCUDA.test_matmul_cuda_float32 that differ from source templat…
PyTorch's test infrastructure dynamically generates test names across devices and dtypes, causing CI failures to show names like TestLinalgCUDA.test_matmul_cuda_float32 that differ from source templat…
Over 100 participants across 20+ teams competed in the ExecuTorch Hackathon in San Francisco on June 27-28, 2026, building on-device AI applications for Snapdragon-powered Samsung Galaxy S25 Ultra dev…
Shopify has joined the PyTorch Foundation as a Platinum member, committing to invest in the open-source AI ecosystem and contribute engineering expertise to shape PyTorch for commerce and agentic AI. …
RadixArk released Miles, an open-source PyTorch-native framework for large-scale LLM reinforcement learning post-training. The framework integrates SGLang for rollout, NVIDIA Megatron-LM for training,…
PyTorch introduced the Cross-Repository CI Relay (CRCR), an automated pipeline that triggers and tracks CI in downstream repositories whenever changes are made to pytorch/pytorch. Results are displaye…
LightSeek Org open-sourced TokenSpeed-Kernel, a portable API and high-performance kernel subsystem for multi-silicon LLM inference, decoupling runtime from hardware-specific code to simplify backend c…
SGLang achieved a 5x throughput improvement for serving DeepSeek-V4 on NVIDIA GB300 hardware since its Day-0 launch, reaching ~11,200 tok/s/GPU at 50 tok/s/user through kernel and runtime optimization…
Meta researchers introduced an LLM-guided autotuner for Helion, PyTorch's domain-specific language for performance portable machine learning kernels, that matches the performance of the existing Likel…
The Linux Foundation Education and PyTorch Foundation have launched the PyTorch Certified Associate (PTCA) certification, validating foundational skills in PyTorch for AI and machine learning practiti…
The schedule for KubeCon + CloudNativeCon + OpenInfra Summit + PyTorch Conference China, taking place September 7-9 in Shanghai, is now available. The event features sessions on AI agents, GPU virtual…
The PyTorch Foundation has opened nominations for its 2026 Contributor Awards, recognizing individuals who strengthen projects like PyTorch, vLLM, DeepSpeed, Ray, Helion, and Safetensors through techn…
Eighty engineers, researchers, and community builders gathered for the inaugural PyTorch Meetup Singapore, hosted at the Red Hat Asia Pacific office. The event featured technical talks on inference, d…
Helion kernels were integrated into vLLM for FP8 inference using Qwen3 models and evaluated across NVIDIA H100 and B200 GPUs. The experiments demonstrated that Helion provides a productive PyTorch-nat…
DeepSpeed has integrated the Muon Optimizer, a memory-efficient optimizer that uses a single momentum buffer and Newton-Schulz orthogonalization to improve training convergence, particularly for 2D we…
The CUDA caching allocator in PyTorch fragments memory when allocated blocks prevent adjacent free blocks from merging, causing allocation failures despite sufficient total free memory. This fragmenta…
LinkedIn re-architected its distributed linear programming solver, DuaLip, using a GPU-accelerated PyTorch version to solve extreme-scale optimization problems involving hundreds of millions of users …
PyTorch's Inductor compiler uses kernel fusion to accelerate model execution by up to 10x, grouping dependent operations into single Triton kernels to reduce memory traffic and kernel launch overhead.…
TokenSpeed, an open-source inference engine, achieved a record-breaking 580 tokens per second running the Qwen3.5-397B-A17B model on GPUs. The performance gain for agentic workloads comes from elimina…
Alibaba Cloud has joined the PyTorch Foundation as a Platinum member, gaining a seat on the foundation's Governing Board and a position on its Technical Advisory Committee. The Chinese cloud computing…
The PyTorch Foundation reopened applications for its Ambassador Program, seeking community leaders to mentor users, create tutorials, and organize events for a two-year term. The foundation especially…