cd/entity/Triton· home› entities› Triton
grep -l @triton /news/*.json | wc -l → 81

Triton

mentions 81 type Organization page 3/5 feed RSS

// recent coverage 81 mentions

13:00
2026-07-29
hiraditya.github.io
machine-learning

When XLA Isn't Enough: Pallas, Mosaic, and Triton

JAX's Pallas kernel system, which lowers through Triton on GPU and Mosaic on TPU, lets developers write custom kernels when XLA's automatic fusion is insufficient for operations like flash attention, …

17:48
2026-07-24
promptcube3.com
artificial-intelligence

Jensen Huang's New X Account: Why It Matters for AI

NVIDIA CEO Jensen Huang's new X account signals a shift toward real-time, decentralized communication with the developer community, potentially accelerating feedback loops on AI hardware and software …

00:00
2026-07-24
blog.getutm.app
developer-tools

Bringup Notes: Building Triton

A developer building the Triton DirectX 11 driver for QEMU argues that AI tools amplify human developers rather than replace them, citing challenges such as sparse documentation, multi-boundary debugg…

06:55
2026-07-23
andytimm.github.io
large-language-models

I audited Stanford's CS336 and built an LLM from scratch for $353

Stanford NLP's CS336 course, audited by a student who built an LLM from scratch for $353, delivers on its premise of deepening understanding of modern LLMs through hands-on assignments that include bu…

17:14
2026-07-22
dev.to
artificial-intelligence

Building Production-Ready RAG Applications: A Practical Guide

A developer's practical guide details the engineering challenges and solutions for deploying production-ready Retrieval-Augmented Generation (RAG) applications, covering data indexing, vector stores, …

19:00
2026-07-21
hiraditya.github.io
machine-learning

How torch.compile Actually Works

Torch.compile is not a traditional compiler but a system that intercepts Python bytecode at runtime, extracts compilable regions through a multi-stage pipeline, and stitches them back with eager Pytho…

03:35
2026-07-21
dev.to
computer-vision

Exploring the Deep Learning Library in Modern Computer Vision

A developer compares PyTorch and TensorFlow for computer vision projects, noting that PyTorch dominates research with dynamic computation graphs and recent performance improvements via torch.compile()…

19:00
2026-07-19
hiraditya.github.io
artificial-intelligence

Triton: The Compiler That Pretends to Be a Library

Triton is a compiler with a Python frontend that parses a function's AST, runs it through an MLIR pipeline, and emits a GPU binary, never executing the Python function as Python. The compiler handles …

02:58
2026-07-16
discuss.huggingface.co
machine-learning

My latest ablation run: integrating Engram onto two backbones

A 200-step, ~1.7B-parameter ablation run in OLMo-core comparing Engram on a standard attention Transformer versus a 3 GDN-layer + 1 attention-layer hybrid found that Transformer + Engram reached sligh…

00:00
2026-07-13
rocm.blogs.amd.com
artificial-intelligence

Triton-Based Optimization of Video Sparse Attention on ROCm

AMD has released a Triton-based optimization for video sparse attention on its ROCm platform, targeting Diffusion Transformers (DiTs) used in video generation. The implementation reduces the quadratic…

00:00
2026-07-13
rocm.blogs.amd.com
artificial-intelligence

GEAK Agent-Driven Optimization of the DeepSeekV4 MLA Kernel

AMD's open-source GEAK agent-driven framework automated the optimization of the DeepSeekV4 MLA kernel, achieving a 2.10x improvement in end-to-end throughput and a 3.71x reduction in time-to-first-tok…

19:12
2026-07-11
byteiota.com
artificial-intelligence

Hugging Face Kernels Are Now Signed Hub Artifacts

Hugging Face announced that custom GPU kernels on its Hub are now signed artifacts governed by a trusted publisher model, requiring a dedicated repository type that replaces the old model-type format.…

16:50
2026-07-09
gimletlabs.ai
ai-agents

Formally Verifying AI-Generated GPU Kernels

Gimlet Labs has built an early research system that uses formal verification to prove semantic equivalence between reference PyTorch models and AI-generated GPU kernels, catching bugs that pass tradit…

← prev page 3 / 5 next →
// co-occurs with top 8 entities
// topics top 6 topics