cd/entity/CUDA· home entities CUDA
grep -l @cuda /news/*.json | wc -l → 319

CUDA

mentions 319 type Organization page 14/16 feed RSS

// recent coverage 319 mentions

20:17
2026-06-16
github.com
artificial-intelligence

Show HN: cuTile Rust: Safe, data-race-free GPU kernels in Rust

NVIDIA Research released cuTile Rust, a tile-based system for writing memory-safe, data-race-free GPU kernels in Rust. The project extends Rust's ownership model to GPU programming, achieving up to 92…

19:37
2026-06-16
dev.to
large-language-models

Serving any LLM using a single command line with Flama

Flama 2.0 introduces first-class support for generative AI, enabling users to download, package, and serve large language models (LLMs) via a single command line. The framework allows fetching models …

17:26
2026-06-16
letsdatascience.com
developer-tools

Lemonade SDK Adds Nvidia CUDA Acceleration Support

Lemonade SDK has added support for Nvidia CUDA acceleration, enabling GPU-accelerated computing for developers using the SDK. The update targets developers working with languages such as C#, Ruby, Pyt…

16:42
2026-06-16
flox.dev
machine-learning

Training NanoGPT on Slurm with a Nix-Pinned Environment

A researcher using nanoGPT on a MacBook Pro faces dependency failures when moving to a GPU cluster, prompting a solution using Nix and Flox to create reproducible, cross-platform runtime environments …

20:40
2026-06-15
gilesthomas.com
machine-learning

Jax: Commitment Issues

JAX's default_device context manager places arrays on the specified device but does not commit them, allowing JAX to move them to other devices. This caused array lookups to take over a second by trig…

15:49
2026-06-15
polarsignals.com
developer-tools

Show HN: Continuous Nvidia CUDA PC Sampling Profiler

Polar Signals has released an open-source, low-overhead continuous profiler for NVIDIA CUDA that supports Program Counter (PC) sampling, allowing developers to see where GPU code spends time at the in…

14:54
2026-06-15
discuss.huggingface.co
large-language-models

CUDA support added - Pre-generation knowledge-boundary estimator

A new pre-generation knowledge-boundary estimator with CUDA support has been developed. The tool uses a small sidecar model and prompt-only features to predict whether a language model will answer cor…

04:00
2026-06-15
arxiv.org
machine-learning

Gefen: Optimized Stochastic Optimizer

Researchers propose Gefen, a memory-efficient optimizer that reduces AdamW's memory footprint by ~8x while maintaining performance, enabling larger microbatches and improved throughput in deep learnin…

18:26
2026-06-14
dev.to
large-language-models

Making a fleet of self-hosted LLM agents trustworthy

LLMKube has introduced a new cluster-scoped CRD, AgentRelease, and a self-update path to make fleets of self-hosted LLM agents trustworthy. The system allows safe, automated updates with atomic symlin…

09:55
2026-06-13
imil.net
large-language-models

RTX 5080 and RTX 3090 Setup: 80 Tok/s on Qwen 3.6 27B Q8

A user combined an RTX 5080 and RTX 3090 on an Asus Prime X570-Pro motherboard to run Qwen 3.6 27B Q8 at over 80 tokens per second. The setup required disabling CSM, enabling Above 4G Decoding and ReS…

12:00
2026-06-12
telnyx.com
large-language-models

Minimax M3 is now available on Telnyx Inference

Telnyx has added MiniMax's M3 model to its Inference platform, offering the first open-weight model combining coding, agent capabilities, and a 1M-token context window. Hosted on Telnyx's B300 GPU inf…

21:40
2026-06-11
docs.rapids.ai
ai-infrastructure

Polars GPU Engine

NVIDIA's cuDF library now provides GPU-accelerated execution engines for the Polars Lazy API, enabling users to run dataframe operations on GPUs with automatic fallback to CPU when operations are unsu…

16:25
2026-06-10
phoronix.com
artificial-intelligence

AMD's Lemonade SDK For Local AI Adds NVIDIA CUDA Support

AMD released Lemonade SDK version 10.7, adding NVIDIA CUDA support to its local AI server solution that previously only supported AMD hardware, Apple Metal GPUs, and AArch64 CPUs. The update integrate…

15:41
2026-06-09
modular.com
ai-infrastructure

What about OpenCL and CUDA C++ alternatives?

Chris Lattner, a lead engineer on Apple's original OpenCL implementation, explains why OpenCL and other C++-based GPU programming models failed to become dominant in AI, citing slow committee-driven d…

← prev page 14 / 16 next →
// co-occurs with top 8 entities
// topics top 6 topics