cd/sources/semianalysis· home› sources› SemiAnalysis
cat /sources/semianalysis.feed | wc -l → 41

SemiAnalysis

articles 41 domain semianalysis.com → page 1/3 feed RSS
13:36
2026-09-26
semianalysis.com
ai-chips

Intel Panther Lake Teardown

SemiAnalysis's STEEL teardown lab found that Intel's Panther Lake, the first commercial chip to use backside power delivery (PowerVia) and Intel's first gate-all-around RibbonFET transistors on the 18…

18:14
2026-09-21
semianalysis.com
ai-infrastructure

Computation and Data Movement for Inference

Mixture of Experts has changed the structure of AI inference serving by altering which tensors are active per token, what must remain close together, which transfers need strong local bandwidth, and h…

18:19
2026-09-13
semianalysis.com
ai-chips

Long Live the Short King: Why 4-hi HBM Wins

Nvidia's next-generation Rubin Ultra accelerator will ship with 192GB of HBM per GPU, down from 288GB in standard Rubin and B300, according to SemiAnalysis, which first reported the change. SemiAnalys…

20:53
2026-09-09
semianalysis.com
artificial-intelligence

Where Does a Robot Think – On-Device vs Datacenter Inference

Physical AI is moving AI beyond screens into robots, but the field is split on where inference should run: on-device versus in datacenters. Figure runs its Helix model entirely onboard, while Physical…

20:00
2026-09-07
semianalysis.com
ai-infrastructure

TPU Inference Externalization Full Steam Ahead - InferenceX

SemiAnalysis published the first third-party inference results for Google's TPUv7 Ironwood, showing up to 50% better performance per dollar than NVIDIA's B200/B300 in apples-to-apples comparisons. The…

20:14
2026-09-01
semianalysis.com
artificial-intelligence

Korea’s Trillion-Dollar Sovereign AI Investment

South Korea's government launched the 'Independent AI Foundation Model' project in June 2025, a tournament-style initiative to develop a domestic frontier AI model, with five consortiums selected in A…

15:46
2026-08-30
semianalysis.com
ai-safety

Most Neoclouds Suck At Security

SemiAnalysis's ClusterMAX 3.0 testing found that most neoclouds have serious security vulnerabilities, with five frightening patterns identified. The company urges neocloud operators and users to upda…

14:00
2026-08-25
semianalysis.com
ai-chips

OpenAI Jalapeño: Better Than Nvidia Blackwell

OpenAI's new inference chip, Jalapeño, outperforms Nvidia's Blackwell and other competitors in performance per watt across almost all scenarios, according to benchmarks run by SemiAnalysis with OpenAI…

16:40
2026-08-21
semianalysis.com
artificial-intelligence

Are Open Models Catching Up?

Open-source AI models are catching up to closed frontier models twice as fast with each generation, according to a SemiAnalysis analysis that found open models now match closed models on coding and ag…

01:32
2026-08-19
semianalysis.com
ai-infrastructure

Cerebras's Next Generation CS-4: Fast Just Got Faster

Cerebras Systems unveiled its fourth-generation CS-4 rack, doubling the performance of the CS-3 by increasing clock speeds and power delivery on the same 5nm wafer-scale engine (WSE-3), enabling doubl…

04:51
2026-08-10
semianalysis.com
artificial-intelligence

Ultra-High Interactivity on NVIDIA GPUs? - TileRT InferenceX

TileRT's persistent engine on NVIDIA GPUs achieves up to 500 tokens/s/user on the InferenceX GLM5 FP8 744B benchmark on a single B200 decode server, approximately 3× faster than GB300 NVL72 running tr…

02:32
2026-08-07
semianalysis.com
artificial-intelligence

Gemini is Cooked but GCP is Cooking

Google announced on August 5th a complete overhaul of DeepMind leadership, with co-founder Demis Hassabis stepping back from day-to-day operations and former Google Chief Scientist Jeff Dean leaving t…

19:42
2026-08-03
semianalysis.com
artificial-intelligence

Kimi K3, The Manos, The Mythos, The Legendos

Moonshot AI released Kimi K3, an open frontier model that swept leaderboards, featuring a hybrid attention mechanism with Kimi Delta Attention (KDA), a linear attention layer derived from DeltaNet and…

22:09
2026-07-29
semianalysis.com
ai-infrastructure

The Wild Wild West Of LEGO Datacenters

Modular construction has become the default playbook for building datacenters, with SemiAnalysis tracking over 61GW of modular capacity across 1,000+ sites and estimating modular penetration will reac…

00:33
2026-07-25
semianalysis.com
artificial-intelligence

Can AMD break the CUDA Moat? AMD Advancing AI 2026

AMD has a great chance of breaking Nvidia's CUDA software moat in AI accelerators, according to SemiAnalysis, which upgraded its outlook from zero to non-zero to now a strong probability of success, c…

page 1 / 3 next →