cd/entity/CUDA· home entities CUDA
grep -l @cuda /news/*.json | wc -l → 318

CUDA

mentions 318 type Organization page 6/16 feed RSS

// recent coverage 318 mentions

00:02
2026-08-10
martinkristiansen.com
machine-learning

Optimizing a GPT-2-Class Transformer on a GPU

A developer's optimization campaign on an RTX 3080 Ti cut a GPT-2-small-class transformer's forward pass from 78.2ms to 1.60ms, a 49× speedup, beating torch.compile's 1.72ms and reaching 136,000 token…

23:47
2026-08-07
gist.github.com
developer-tools

Open WebUI + Ollama + Hugging Face on macOS and Windows

A developer has published a practical guide for setting up a local AI troubleshooting and support environment on macOS and Windows using Open WebUI, Ollama, and Hugging Face. The recommended architect…

12:25
2026-08-06
techpowerup.com
artificial-intelligence

NVIDIA Neural Texture Compression Now Runs on RTX Spark

NVIDIA has brought its RTX Neural Texture Compression (NTC) technology to Windows-on-Arm, ahead of the general release of its RTX Spark PC platform. NTC, first demonstrated at GTC 2026 in May, can red…

01:05
2026-08-06
hazyresearch.stanford.edu
artificial-intelligence

Retire the Abstractions

Hazy Research, the team behind ThunderKittens and Megakernels, argues that CUDA domain-specific languages (DSLs) are heading toward retirement as AI agents take over the cognitive offloading role that…

11:48
2026-08-05
gist.github.com
machine-learning

Cloud Training on RunPod: A Field Guide to the Edge Cases

AlphaPebble Labs engineers detailed a field guide for training AI models on RunPod's rented GPU infrastructure, highlighting edge cases such as the SSH gateway acting as a console rather than an exec …

01:06
2026-08-05
letsdatascience.com
ai-infrastructure

AMD Pitches Open ROCm as Its Counterweight to Nvidia’s CUDA

AMD reported record second-quarter revenue of $11.5 billion, up 50% year over year, with data center revenue reaching $6.7 billion, up 107%, as CEO Lisa Su said open-source contributions to its ROCm s…

10:58
2026-08-04
discuss.huggingface.co
artificial-intelligence

Decay-Gated O(N) Causal Linear Attention with Fused Triton Kernel

A developer has open-sourced a Decay-Gated O(N) Causal Linear Attention architecture with fused Triton/CUDA kernels, aiming to bypass quadratic multi-head attention bottlenecks. The project includes a…

03:25
2026-08-04
promptcube3.com
artificial-intelligence

CUDA's Moat Is Weakening, and AI Coding Agents Are the Pickaxe

AI coding agents are eroding Nvidia's CUDA software moat by lowering the expertise barrier for GPU kernel development, according to a tech analysis. Tools like Anthropic's Claude Code can generate per…

00:26
2026-08-04
promptcube3.com
artificial-intelligence

US vs China AI: the lead is basically gone

The US lead over China in AI has essentially disappeared, according to an analysis of model releases and deployment trends. Chinese models like DeepSeek's R1 and V3 and Qwen now match or beat US open-…

03:28
2026-08-02
gist.github.com
developer-tools

cuda-oxide: eliding bounds checks with proof carrying views

The cuda-oxide project introduces proof-carrying views that eliminate bounds-check overhead in CUDA kernels written in Rust, boosting GEMM performance from 2,942 to 7,159 GFLOPS (2.43x) with only ~0.1…

19:10
2026-07-31
promptcube3.com
ai-chips

Hygon's 512-Thread CPU and AI GPU: Intel/Nvidia Rival?

Hygon Information Technology Co., Ltd. has unveiled a 512-thread x86 CPU and an AI GPU, positioning itself as a potential rival to Intel Corporation's Xeon processors and Nvidia Corporation's datacent…

03:20
2026-07-31
dev.to
developer-tools

Why NVIDIA Open-Sourced Its Linux GPU Kernel Modules

NVIDIA has open-sourced its Linux GPU kernel modules, including nvidia.ko, nvidia-drm.ko, nvidia-uvm.ko, and nvidia-modeset.ko, while keeping user-space components like CUDA and GSP firmware proprieta…

← prev page 6 / 16 next →
// co-occurs with top 8 entities
// topics top 6 topics