State of Open Models: Summer 2026 Observations
Hugging Face's biannual report on open models for summer 2026 finds that Chinese labs released the largest open models in almost every month of 2026, with monthly ceilings ranging from 754B to 2.78 tr…
Hugging Face's biannual report on open models for summer 2026 finds that Chinese labs released the largest open models in almost every month of 2026, with monthly ceilings ranging from 754B to 2.78 tr…
Dell's Pro Precision 5 14s AMD mobile workstation, powered by the 12-core, 24-thread AMD Ryzen AI 9 HX PRO 475 processor with a 60 TOPS NPU and Radeon 890M integrated graphics, starts at $2,253 and is…
OpenAI acquired a 4.22% economic stake in Cerebras Systems in July by exercising warrants for 10,033,508 Class N shares at $0.00001 each, a cash cost of about $100.34, ahead of the launch of a Cerebra…
GitHub has launched the session catalog for GitHub Universe 2026, a two-day event featuring sessions from AMD, Figma, NVIDIA, Coinbase, Anthropic, and OpenAI, with early bird registration saving $300 …
A new guide outlines strategies for scaling AI revenue while controlling cloud costs, covering hardware accelerators from NVIDIA, Groq, and Cerebras, inference engines like vLLM and TensorRT-LLM, and …
Corsair's Vengeance RGB RS DDR5 16GB (2x8GB) 5200MHz CL40 RAM kit is available on Amazon at 18% off recent price hikes, offering a rare saving amid record-high memory prices driven by AI demand. The k…
A software developer is building a home AI data center from e-waste, using four AMD V620 GPUs (32GB VRAM each) purchased cheaply from eBay resellers, an Intel Core i9 10900X on an X299 motherboard, an…
Samsung Electronics Co. has sold out its entire 2026 supply of HBM4 memory chips, with customer demand concentrated on the product, according to Chief Financial Officer Park Soon-cheol on the company'…
Cerebras Systems CEO Andrew Feldman announced a partnership with AMD on July 23, combining Cerebras' Wafer-Scale Engine with AMD's Helios rack systems to create a disaggregated inference solution targ…
AMD and PyTorch upstreamed FP8 training optimizations for AMD Instinct GPUs into TorchAO and TorchTitan, delivering a 13.4% throughput gain over BF16 on Llama3-8B dense models and recovering 89% of FP…
AMD released MiniDXNN v0.4.0, an open-source library for GPU-accelerated MLP inference and training on DirectX 12, adding an interactive GUI application for neural texture compression. The update incl…
Riot Platforms sold 4,300 BTC in Q2 2026, reducing its treasury from 15,680 to 11,380 BTC, to fund operations and its pivot toward AI infrastructure. The NASDAQ-listed Bitcoin miner reported revenue o…
VectorWare has enabled Rust's portable SIMD (core::simd) to run natively on GPUs, allowing the same SIMD code to execute on x86 CPUs, ARM, and NVIDIA GPUs without modification. This breakthrough maps …
LG's most affordable 2026 B6 OLED TVs are discounted by up to $800, with prices starting at $1,700 for the 65-inch model, and include a free $100 or $200 Fanatics gift card. The 77-inch model is $2,50…
Anthropic signed a 20-year, $9.1 billion compute agreement with Riot Platforms on August 11, converting the former Bitcoin miner's 700 MW Rockdale, Texas campus into AI infrastructure. This is Anthrop…
Luminance, a legal AI platform founded in 2015 out of Cambridge mathematics research, has raised roughly $165 million across seven rounds, including a $75 million round to expand its legal agents, and…
Liqid unveiled its UltraStack 30 platform, pooling up to 30 AMD Instinct MI350P GPUs in a single AMD EPYC-based server, with 4.3 TB of aggregate HBM3E memory and up to 69 PFLOPS of FP8 performance. Th…
Liqid unveiled the UltraStack 30, a scale-up AI platform that pools up to 30 AMD Instinct MI350P GPUs in a single server, delivering 4.3TB of HBM3E memory and 69 PFLOPS of FP8 compute for AI inference…
AMD's Instinct MI455X accelerator, unveiled at the 2026 Advancing AI event, introduces the CDNA 5 architecture, delivering 4x peak FP4/FP8 matrix performance and 2x other compute versus the MI355X, bu…
AMD Corporate VP Madhu Rangarajan said there is no one-size-fits-all for agentic AI infrastructure, arguing that workloads labeled as such vary widely in memory, storage, and CPU needs. He noted that …