cd/entity/Hazy Research· home entities Hazy Research
grep -l @hazy research /news/*.json | wc -l → 5

Hazy Research

mentions 5 type Person feed RSS

// recent coverage 5 mentions

09:00
2026-09-08
cohere.com
artificial-intelligence

Inside the megakernel serving engine for North Mini Code

Cohere released a serving engine for its North Mini Code model built around a decode megakernel that runs 1.25x to 1.41x faster than vLLM end-to-end on a single H100 with BF16 precision. The engine, a…

22:09
2026-08-16
cryptobriefing.com
artificial-intelligence

AI efficiency jumped 18x in 16 months, Stanford research finds

Stanford's Hazy Research group found that the intelligence per joule (IPJ) of local AI inference improved 18-fold from mid-2024 to late 2025, driven by a 3.1x gain from model architecture improvements…

14:50
2026-08-14
techcrunch.com
artificial-intelligence

Kog is going deeper to squeeze more inference out of GPUs

French startup Kog claims its software can deliver 30x faster LLM inference on standard datacenter GPUs, demoing 3,000 tokens per second on a 2-billion-parameter model using AMD MI300X and NVIDIA H200…

01:05
2026-08-06
hazyresearch.stanford.edu
artificial-intelligence

Retire the Abstractions

Hazy Research, the team behind ThunderKittens and Megakernels, argues that CUDA domain-specific languages (DSLs) are heading toward retirement as AI agents take over the cognitive offloading role that…

07:53
2026-06-18
letsdatascience.com
artificial-intelligence

Study Finds Small Desktop AI Challenges Data-center Models

A Stanford University study published November 2025 introduced "intelligence per watt" (IPW) as a metric for AI inference efficiency, finding that local language models can accurately answer 88.7% of …

// co-occurs with top 8 entities
// topics top 6 topics