cd/entity/TPU· home entities TPU
grep -l @tpu /news/*.json | wc -l → 56

TPU

mentions 56 type Organization page 1/3 feed RSS

// recent coverage 56 mentions

02:30
2026-09-07
aiflash.com
artificial-intelligence

MaxKernel: Agentic Kernel Generation for TPUs

MaxKernel, a new agentic system presented by Google, uses large language models with real-time compiler feedback to generate high-performance custom kernels for Tensor Processing Units (TPUs), address…

17:47
2026-08-30
dev.to
artificial-intelligence

CPU, GPU, TPU, NPU, DPU, QPU: six chips, one question

Maneshwar, the developer behind LiveReview, an AI code review tool, explains the differences between CPU, GPU, TPU, NPU, DPU, and QPU, arguing that 'computation' is not a single thing. He details how …

19:07
2026-07-31
promptcube3.com
ai-infrastructure

AI's Biggest Spenders Are Accelerating — Capex Deep Dive

Microsoft, Google, Amazon, and Meta are accelerating capital expenditures on AI infrastructure, with Microsoft expanding Azure data center leases and GPU fleets, Google committing to TPU buildouts, Am…

21:04
2026-07-30
ycrootaccess.com
artificial-intelligence

Jeff Dean: The 1% Rule for Building in AI

Google Chief Scientist Jeff Dean said at Y Combinator's Startup School 2026 that AI models have reached the capability of junior engineers, a prediction he made in May 2025, and predicted that by 2027…

16:21
2026-07-30
developers.googleblog.com
artificial-intelligence

How to use Google microbenchmarks for evaluating TPU performance

Google released a microbenchmark suite on GitHub to help developers evaluate TPU performance by measuring network, compute, HBM, and host transfer capabilities. The suite establishes a Speed-of-Light …

19:05
2026-07-25
promptcube3.com
machine-learning

MoE Capacity Factor: Why Your Tokens are Being Dropped

Token dropping in mixture-of-experts (MoE) layers occurs when a capacity factor limits each expert's buffer, causing excess tokens to bypass the expert MLP and degrade model quality under production l…

19:00
2026-07-25
hiraditya.github.io
machine-learning

XLA Up Close: What It Optimizes, and What It Won't

XLA, the compiler behind JAX, TensorFlow, and PyTorch/XLA, optimizes array programs by freezing shapes, statically allocating buffers, and fusing operations against a global cost model, which makes it…

21:13
2026-07-23
thedailycompute.beehiiv.com
artificial-intelligence

Google reportedly working on ultra-efficient AI chip for Gemini

Google is reportedly developing 'Frozen v2,' a dedicated inference chip built specifically to run Gemini models, claiming a 6x to 10x increase in energy efficiency by hardcoding core Gemini operationa…

14:45
2026-07-23
i-programmer.info
artificial-intelligence

Google Releases AlphaEvolve On Gemini

Google has released AlphaEvolve, an AI coding agent developed by Google DeepMind, on the Gemini Enterprise Agent Platform. The tool uses large language models to automatically generate and improve alg…

13:49
2026-07-23
promptcube3.com
ai-chips

Nvidia Alternatives: The Surge in AI Chip Demand

A surge in demand for non-Nvidia AI chips is reshaping hardware strategies as engineers optimize large language model deployments for alternative architectures due to H100 shortages and high costs. Th…

page 1 / 3 next →
// co-occurs with top 8 entities
// topics top 6 topics