cd/entity/JAX· home entities JAX
grep -l @jax /news/*.json | wc -l → 81

JAX

mentions 81 type Organization page 1/5 feed RSS

// recent coverage 81 mentions

16:09
2026-09-09
sourcefeed.dev
artificial-intelligence

Ray's new TPU support is aimed at your GPU bill

Ray 2.55 adds official TPU support with atomic gang scheduling for TPU slices, aiming to cut GPU costs by enabling Ray-based workloads to run on Google's TPUs. Google's motive is to boost external TPU…

22:46
2026-09-07
frontierroles.com
artificial-intelligence

Thermo ML Resident — Extropic

Extropic, a Boston-based hardware startup, is hiring junior ML scientists for its Thermo ML Resident program, offering a salary of $75,000–$200,000 per year for on-site work. Residents will collaborat…

21:21
2026-09-03
frontierroles.com
artificial-intelligence

Staff VLA Engineer — 42dot

42dot, the Global Software Center of Hyundai Motor Group, is hiring a Staff VLA Engineer in Sunnyvale, California, with a salary range of $189,000–$311,000 per year, to research and prototype next-gen…

17:25
2026-09-03
gilesthomas.com
artificial-intelligence

Putting my JAX-trained models on the Hugging Face Hub

Developer gpjt has uploaded PyTorch-compatible versions of all his JAX-trained GPT-2 models to the Hugging Face Hub, including models from his blog series on writing an LLM from scratch and Chinchilla…

00:38
2026-08-29
dev.to
artificial-intelligence

Pure JAX on G5g: Serving Gemma 4 on Graviton and a T4G

A developer has published a step-by-step guide for serving Google's Gemma 4 model on an AWS EC2 G5g instance using pure JAX, targeting the cheapest whole NVIDIA GPU available on AWS. The guide details…

22:33
2026-08-27
frontierroles.com
ai-research

Research, Coding Agents — Thinking Machines Lab

Thinking Machines Lab, a San Francisco-based AI company, is hiring a Coding Agents researcher with a salary range of $350k–475k per year to work on agentic coding capabilities for its frontier models.…

08:58
2026-08-27
trace.vladsavinov.com
ai-infrastructure

Show HN: Reconstruct distributed LLM training traces

A developer has released a tool to reconstruct and visualize distributed LLM training traces, enabling users to drag and drop compute kernels and collectives to create DDP, TP, FSDP, EP, and CP traces…

17:09
2026-08-26
sourcefeed.dev
artificial-intelligence

Google's 4.7x Qwen 3.5 Speedup Is a Sharding Story

Google engineers reported a 4.7x faster prefill and 3.1x faster decode for Qwen 3.5-397B-A17B on Ironwood TPUs between April and June, achieved by using data parallelism for attention layers and exper…

01:12
2026-08-20
gilesthomas.com
machine-learning

Use the built-in GELU, don't roll your own!

PyTorch's built-in GELU function is 20% faster than a hand-rolled version when training GPT-2 small models, according to a developer's benchmark. The same code training the same model on the same data…

23:09
2026-08-19
frontierroles.com
artificial-intelligence

Software Engineer, Trainium — OpenAI

OpenAI is hiring a Software Engineer, Trainium in San Francisco with a salary range of $295k–380k/yr, to build and optimize its inference stack for AWS Trainium, developing high-performance kernels an…

19:55
2026-08-18
ezyang.github.io
machine-learning

How to Parallelize a Transformer for Training

A new interactive adaptation of the JAX scaling book chapter on transformer training parallelism lets readers drag sliders to see how data, tensor, pipeline, and expert parallelism affect communicatio…

page 1 / 5 next →
// co-occurs with top 8 entities
// topics top 6 topics