Modular: Qualcomm to Acquire Modular
Qualcomm announced an agreement to acquire Modular Inc, an AI-native software platform company, to strengthen its software foundation for generative and agentic AI across data center and edge environm…
Qualcomm announced an agreement to acquire Modular Inc, an AI-native software platform company, to strengthen its software foundation for generative and agentic AI across data center and edge environm…
Modular released version 26.4 of its platform, adding state-of-the-art mixture-of-experts (MoE) serving on Modular Cloud, support for new open-weight models like MiniMax M3 and GLM 5.2, and advancing …
Modular announced version 26.4 of its platform, featuring state-of-the-art Mixture-of-Experts serving, model bringup via agent skills, and the release of Mojo 1.0 Beta 2.…
Modular announced its annual developer conference ModCon 2026, themed 'Compute Unlocked,' will take place on August 18th in San Francisco. The event focuses on solving hardware scarcity by enabling AI…
MiniMax released the open-weights MiniMax M3 model on Modular Cloud, featuring a new Sparse Attention operation that achieves up to 15.6x speedup on decode while maintaining a 1 million token context …
Modular has introduced a five-stage composable routing system for large language model inference, replacing traditional fixed algorithms like round-robin and consistent hashing. The system, detailed i…
At MLSys 2026, Modular identified three key trends in AI inference, highlighted by keynotes on agentic kernel development and the need for "zero trust" verification to prevent benchmark cheating. Lido…
Modular released the first beta of Mojo 1.0, a systems programming language that combines Python-like syntax with Rust-style memory safety and ownership. The language provides precise control over mem…
Modular has built a new data layer for LLM inference routing that solves the problem of querying cached blocks across hundreds of pods in microseconds. The company's architecture uses a specialized da…
A developer built a production pastebin service called mobin using only the Mojo programming language, creating 10 supporting libraries and a documentation tool from scratch in a few weeks with AI cod…
Hippocratic AI partnered with Modular to integrate the MAX framework into its inference pipelines, achieving sub-500ms mean time to first token and approximately 30% faster P99 end-to-end latency for …
Modular released AI agent skills for its Mojo language that enable coding assistants to translate existing GPU kernels from CUDA and Triton into Mojo code, addressing the challenge that large language…
Modular CEO and Google AI infrastructure veteran Chris Lattner launched Inkwell, a real-time illustrated storybook app built on the company's Modular Cloud inference platform, demonstrating sub-second…
Modular announced that traditional HTTP-era load balancing algorithms like round-robin, consistent hashing, and least-connections are inadequate for large language model inference because GPU pods are…
Modular hosted an AMD AI DevDay event and opened new offices, while its community continued to ship new projects and contributions. The company's latest Modverse update highlights ongoing collaboratio…