cd/entity/B200· home entities B200
grep -l @b200 /news/*.json | wc -l → 64

B200

mentions 64 type Organization page 2/4 feed RSS

// recent coverage 64 mentions

05:43
2026-08-11
promptcube3.com
ai-infrastructure

Nvidia chips are basically the new digital gold according to

Nvidia's H100 and B200 chips are increasingly treated as appreciating assets rather than depreciating capital expenditures, according to a commentary on the AI hardware market. This shift is driven by…

23:41
2026-08-10
seangoedecke.com
artificial-intelligence

No, local models will not win

Local AI models will not win because datacenter inference is inherently cheaper and more powerful, according to an analysis by an unnamed author. The author argues that datacenter models benefit from …

21:28
2026-08-10
promptcube3.com
ai-infrastructure

Nvidia is basically forcing Wall Street to fund the AI

Nvidia is shifting AI infrastructure financing from operational expenditure to capital expenditure, positioning itself as the anchor that makes these investments safe for Wall Street, according to a n…

19:02
2026-08-05
tokenstead.ai
large-language-models

Pokee-Isaac 28B

Pokee AI released Pokee-Isaac 28B, a 28B-parameter proprietary non-decoder-only model claiming a 10M-token context that fits on a single RTX 4090 (24GB) in quantized form. Vendor-reported benchmarks i…

16:15
2026-08-05
twitter.com
artificial-intelligence

Pokee-Isaac 28B: 10M-token context, deployable on a single GPU

Pokee-Isaac 28B, the world's first real 10M-token context frontier-class agentic model, is now deployable on a single GPU starting from RTX 4090, according to its release announcement. The model achie…

20:34
2026-08-03
twitter.com
artificial-intelligence

DeepSeek-V4-Flash 2.98x faster on 4x B200, lossless

RunInfra AI optimized DeepSeek-V4-Flash-0731, achieving a 2.98x speedup on 4x B200 GPUs, with median latency reduced from 9248 ms to 3095 ms and throughput increased from 113 to 363 tokens per second,…

11:00
2026-07-30
dev.to
artificial-intelligence

Atomarine: Nuclear Data Centers at Sea!

Atomarine proposes deploying small modular nuclear reactors on maritime vessels to power floating data centers, addressing energy bottlenecks for AI training clusters. A 4,000-GPU cluster requiring 40…

04:43
2026-07-29
promptcube3.com
artificial-intelligence

Chip Stocks: Why the AI Hype Cycle is Hitting a Wall

Semiconductor stocks are sliding across US and Asian markets as investors shift from the AI hype cycle to demanding proof of returns on massive capital expenditures. Big Tech is spending billions on N…

00:12
2026-07-29
promptcube3.com
artificial-intelligence

Apple's $5 Trillion Milestone vs the AI Stock Rotation

Apple Inc. reached a $5 trillion market capitalization milestone as investors rotate away from AI pure-play stocks, reflecting a market shift from AI training to inference and deployment. The move tow…

00:00
2026-07-29
cefboud.com
large-language-models

How Profitable is LLM Inference? Doing the Math on Kimi K3

LLM inference profitability depends on the trade-off between batch size and GPU count, which determines token latency and cost per million tokens. Applying this model to Kimi K3, which requires at lea…

21:57
2026-07-28
promptcube3.com
ai-chips

Nvidia's Market Strategy

Nvidia's market strategy faces a hardware bottleneck as government trade restrictions force the company to create 'lite' versions of its H100 and B200 chips, such as the H20, to meet legal requirement…

20:05
2026-07-24
promptcube3.com
artificial-intelligence

Google's New AI Chip: Reducing Gemini's Inference Costs

Google has developed a new custom AI chip designed specifically for its Gemini large language model, aiming to drastically reduce inference costs and latency by addressing the memory wall and power in…

← prev page 2 / 4 next →
// co-occurs with top 8 entities
// topics top 6 topics