NVIDIA Vera targets the agent-loop bottleneck
NVIDIA published a new CPU category on 7 July with Vera, an Arm server CPU designed to address the agentic-AI bottleneck by maximizing single-threaded performance rather than core density. In testing …
NVIDIA published a new CPU category on 7 July with Vera, an Arm server CPU designed to address the agentic-AI bottleneck by maximizing single-threaded performance rather than core density. In testing …
AMD has integrated an NVFP4 emulation pipeline into vLLM that enables AMD Instinct MI355 accelerators to serve standard NVFP4 quantized checkpoints directly, dequantizing weights to BF16 on-the-fly at…
AMD's QuickReduce library now supports INT3 quantization for all-reduce communication in multi-GPU LLM inference, achieving a 22% reduction in on-wire data volume compared to INT4 on AMD Instinct MI35…
AMD's open-source GEAK agent-driven framework automated the optimization of the DeepSeekV4 MLA kernel, achieving a 2.10x improvement in end-to-end throughput and a 3.71x reduction in time-to-first-tok…
Meta plans to begin production of its new AI chip, Iris, in September, aiming to reduce reliance on Nvidia GPUs. The company targets 14 gigawatts of AI computing capacity by 2027 and has raised capita…
SK Hynix debuted on the Nasdaq with a $26.5 billion share sale, the second-largest in U.S. history, as shares rose 12.8% on the first day. The listing gives U.S. investors direct access to the leading…
The US reclassified the UAE to Country Group A:5, effective July 10, 2026, eliminating export licenses for advanced AI chips, military gear, and satellites. The move follows the US-UAE AI Cooperation …
ZML released LLMD, a free inference server for open large language models that runs across Nvidia CUDA, AMD ROCm, Google TPU, Intel oneAPI and Apple Metal, aiming to decouple AI workloads from proprie…
Altera, the world's largest pure-play FPGA provider, returned to growth with roughly 20% annual revenue growth and more than doubled operating income, driven by demand for its programmable chips in AI…
Meta is transitioning from an ad-driven giant to a vertically integrated AI hyperscaler with its custom 'Iris' AI chip scheduled for production in September 2026 and the launch of 'Meta Compute,' a ne…
Altera, the programmable chip maker spun out of Intel, is growing roughly 20% annually and more than doubling operating income as it prepares for an eventual public listing, CEO Raghib Hussain said. T…
Mini PCs with unified memory architectures, such as those using AMD's Strix Halo chip, can run 70-billion-parameter language models that exceed the VRAM capacity of discrete GPUs like the RTX 5090, bu…
Japan-backed Rapidus aims to mass-produce 2 nm chips by 2027 at lower prices than TSMC, but TSMC's dominant market share, customer relationships, and manufacturing scale make it difficult for challeng…
Goldman Sachs raised AMD's price target to $640, citing surging demand for AI chips driven by agentic AI systems. AMD stock has gained over 100% year-to-date as analysts project sold-out capacity and …
The US Commerce Department reclassified the United Arab Emirates from export control groups D:3 and D:4 to Group A:5, allowing license-free access to advanced AI chips from Nvidia and AMD. The move, t…
AMD's Ryzen AI Halo processor delivers strong performance in AI workloads, according to a hands-on review from Micro Center. The chip integrates a dedicated neural processing unit for on-device machin…
Meta plans to begin production of its proprietary 'Iris' AI chip in September 2026 as part of a strategy to double computing power to 14 gigawatts by 2027 and deliver personal super intelligence to us…
Best Buy is offering Lenovo's 2026 Yoga 7a 2-in-1 touchscreen OLED Copilot+ PC for $749.99, down from $1,200, as part of its Back to School sale. The deal marks a new low price for the laptop, which f…
Supercomputing is splitting into two tracks in 2026: publicly funded exascale systems ranked by TOP500 FLOPS and hyperscaler-built AI campuses measured in megawatt–gigawatt capacity. Microsoft claimed…
A tech worker reflects on the rapid advancement of AI hardware, from childhood PCs to a modern multi-GPU system with 2.5TB of VRAM, while expressing deep skepticism about the AI industry's hype, ethic…