cd/entity/NVIDIA H100· home entities NVIDIA H100
grep -l @nvidia h100 /news/*.json | wc -l → 36

NVIDIA H100

mentions 36 type Person page 1/2 feed RSS

// recent coverage 36 mentions

15:38
2026-08-27
promptcube3.com
artificial-intelligence

Why hasn't the military spearheaded the current AI revolution?

The military has not spearheaded the current AI revolution because modern LLM development depends on vast, unclassified civilian data, fast commercial capital cycles, and the dual-use nature of the te…

00:00
2026-08-27
mindstudio.ai
artificial-intelligence

Breeze TTS 2: Specs, VRAM Needs, and Local Setup Guide

BreezeBlue released Breeze TTS 2, an open-weight bilingual text-to-speech model with sub-40ms time to first audio on an NVIDIA H100, 0.32 real-time factor streaming, and voice cloning, design, and dir…

00:25
2026-08-23
github.com
artificial-intelligence

NanoGPT Speedrun

The NanoGPT speedrun repository reports that a collaborative effort has trained a language model to 3.28 cross-entropy loss on the FineWeb validation set in under 75 seconds on 8 NVIDIA H100 GPUs, a d…

17:36
2026-08-21
liquid.ai
artificial-intelligence

LFM2.5-DSpark: Up to 3.2x Faster Inference from H100 to MacB

Liquid AI released DSpark draft model checkpoints for three LFM2.5 models, enabling speculative decoding that speeds up inference by up to 3.18x on GPUs and 2.87x on-device without changing output qua…

04:00
2026-08-19
arxiv.org
artificial-intelligence

KernelArc: A Multi-Agent Framework for GPU Kernel Optimization

KernelArc, a multi-agent framework for autonomous GPU kernel optimization, achieved first-place rankings on representative L1, L2, Quantization, and FlashInfer tasks at the public SOL-ExecBench leader…

21:57
2026-08-16
promptcube3.com
artificial-intelligence

AI debt bubbles are going to force the Fed's hand again

A looming AI debt bubble, driven by massive infrastructure spending on GPU clusters and data centers, threatens systemic financial stability, according to analysis of the tech sector's capital expendi…

03:08
2026-08-16
sourcefeed.dev
artificial-intelligence

The Real Cost of AI-Porting 250k Lines of Fortran

A team from Nagoya University and the University of Tokyo used Claude Code to port CReSS, a 250,000-line Fortran typhoon simulator, to GPUs, achieving 162 validated GPU kernels and a 5.1x speedup on a…

01:03
2026-08-16
changyi.fun
machine-learning

Attention Through Arithmetic Intensity

A technical analysis by an unnamed author, inspired by a post from Su Jianlin, derives the arithmetic intensity (FLOPs per byte) of attention during single-token decode, finding MHA has an AI of exact…

17:58
2026-08-11
attestable.com
artificial-intelligence

Proving Muse Glimmer in Zero-Knowledge at 85 Tok/S on an H100

Attestable, a zero-knowledge proof startup, announced it can prove production-scale transformer inference at 85 tokens per second on a single NVIDIA H100, with proof sizes of 4.35 to 7.92 MiB and veri…

13:30
2026-08-10
cast.ai
artificial-intelligence

LLM Inference Cost Optimization: Run AI Inference for Less

Cast AI benchmark testing shows that continuous batching at batch size 8 reduces Llama 3.1 70B inference cost on a single H100 from approximately $0.60-$0.80 per million tokens to $0.15-$0.25 per mill…

16:00
2026-08-06
cloud.google.com
artificial-intelligence

Advancing brain tumor research with privacy-first AI

Google Cloud is collaborating with MLCommons through the MedPerf initiative to use Confidential Computing for privacy-first benchmarking of medical AI models, enabling validation on diverse real-world…

04:00
2026-08-04
arxiv.org
large-language-models

DiffusionGemma Technical Report

Google DeepMind introduced DiffusionGemma, an experimental open-weight language model that uses discrete diffusion to generate text at high speed, achieving roughly 1,500 output tokens per second on a…

22:31
2026-07-27
aneforge.com
artificial-intelligence

ANEForge: Program the Apple Neural Engine Directly, from Python

ANEForge, a new open-source Python package, compiles tensor graphs directly into programs for the Apple Neural Engine (ANE) and runs them without CoreML, enabling both inference and training on the ch…

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics