cd/entity/NVIDIA H100· home entities NVIDIA H100
grep -l @nvidia h100 /news/*.json | wc -l → 36

NVIDIA H100

mentions 36 type Person page 2/2 feed RSS

// recent coverage 36 mentions

00:00
2026-07-23
fergusfinn.com
ai-infrastructure

Throughputmaxxing DeepSeek-V4-Flash on Isambard-AI

Doubleword, one of six companies in the first wave of UK Sovereign AI investments, achieved up to 3× the throughput of vLLM for DeepSeek-V4-Flash on a single node of Isambard-AI, the UK's national AI …

15:02
2026-07-22
dev.to
artificial-intelligence

Google's Gemma 2 is here. It's a big deal for open models.

Google has released Gemma 2, the next generation of its open models, featuring a 27B parameter version that offers performance competitive with models more than twice its size while being efficient en…

19:12
2026-07-17
insideai.news
artificial-intelligence

Meta in Talks for $10 Billion Anthropic Compute Deal in the US

Meta Platforms is in advanced talks to lease computing power to Anthropic in a deal that could be worth $10 billion, according to a New York Times report. The agreement would make Meta a critical infr…

15:02
2026-07-08
dev.to
large-language-models

Gemma 2 is here. The architectural tweaks are what matter.

Google released Gemma 2, an open model family with 9B and 27B parameter sizes, featuring architectural changes like hybrid attention and Grouped-Query Attention for improved inference efficiency. The …

18:02
2026-06-23
discuss.huggingface.co
machine-learning

Looking for arXiv cs.LG endorser for PsiLogic optimizer

Ali, a 16-year-old independent researcher, developed PsiLogic, a chaos-aware optimizer built on Adam, and seeks an arXiv cs.LG endorser to submit his preprint. His benchmarks on an NVIDIA H100 GPU sho…

15:02
2026-06-19
dev.to
large-language-models

Gemma 2's Architecture: More Performance from Less Model

Google's Gemma 2 models demonstrate that architectural efficiency can deliver competitive performance with fewer parameters. The 27B model rivals models twice its size through hybrid attention, Groupe…

16:36
2026-06-15
recursive.com
artificial-intelligence

First Steps Toward Automated AI Research

Recursive's automated AI research system achieved state-of-the-art results on three benchmarks: fixed-budget language model training, small-model training speed, and GPU kernel optimization. The syste…

14:01
2026-06-15
dev.to
large-language-models

Fine-Tune Llama 3 706B Model Locally

Nick Creighton, an operator who ships, provides a detailed blueprint for deploying Meta's Llama 3 706B model locally, emphasizing privacy, latency, and cost benefits over cloud APIs. He outlines the e…

01:13
2026-06-14
byteiota.com
large-language-models

DiffusionGemma: Google’s 4x Faster Text Diffusion Model

Google DeepMind released DiffusionGemma on June 10, 2026, a 26B open-weight text diffusion model that generates 256 tokens simultaneously, achieving up to 1,008 tokens per second on an H100—4-5x faste…

← prev page 2 / 2
// co-occurs with top 8 entities
// topics top 6 topics