cd/entity/CUDA· home entities CUDA
grep -l @cuda /news/*.json | wc -l → 319

CUDA

mentions 319 type Organization page 10/16 feed RSS

// recent coverage 319 mentions

21:06
2026-07-13
sourcefeed.dev
artificial-intelligence

Discarded Teslas Still Deliver Local AI VRAM

Discarded NVIDIA enterprise Tesla GPUs, including K80 (24 GB, ~$60), P100 (16 GB, ~$75), and V100 (16 GB, under $200), remain usable for local AI inference and training workloads when paired with olde…

16:53
2026-07-13
tornadovm.org
developer-tools

What If Java Apps Could Access CUDA Ecosystem Gracefully

TornadoVM now natively integrates NVIDIA CUDA libraries cuBLAS, cuFFT, and cuDNN, allowing Java applications to call GPU-accelerated linear algebra and deep learning primitives directly from JIT-compi…

12:22
2026-07-12
blawg.pages.dev
artificial-intelligence

Ollama vs. Llama.cpp – Quick Benchmark

A benchmark comparing Ollama and llama-server on a Tesla V100 found that both backends achieve nearly identical token generation speeds (about 111 tok/s), but llama-server processed prompts 22% faster…

00:04
2026-07-12
sourcefeed.dev
large-language-models

Fine-Tune Qwen2.5-7B with QLoRA on Your Own Data

Mariana Souza published a practical guide for fine-tuning Qwen2.5-7B-Instruct using QLoRA on custom instruction datasets, including cost estimates and a loss-masking sanity check. The tutorial covers …

19:12
2026-07-11
byteiota.com
artificial-intelligence

Hugging Face Kernels Are Now Signed Hub Artifacts

Hugging Face announced that custom GPU kernels on its Hub are now signed artifacts governed by a trusted publisher model, requiring a dedicated repository type that replaces the old model-type format.…

02:39
2026-07-11
machinebrief.com
machine-learning

Breaking Graph Bottlenecks: GPU Power Unleashed

Researchers have developed a new approach to 1-WL stable coloring for Graph Neural Networks that breaks traditional scalability bottlenecks by using a probabilistically backed refinement algorithm. A …

15:47
2026-07-10
github.com
artificial-intelligence

Show HN: Real-time n-body tree code in CUDA

A developer released a real-time n-body simulation using the Barnes-Hut algorithm on GPU, implemented in CUDA C++ and OpenGL, scaling to millions of particles on an NVIDIA RTX 500 Ada Laptop GPU. The …

13:24
2026-07-10
arxiv.org
large-language-models

DominoTree

Researchers introduced DominoTree, a training-free best-first draft tree method for speculative decoding that uses Domino's conditional correction to achieve up to 6.6x speedup over autoregressive dec…

09:51
2026-07-09
glukhov.org
artificial-intelligence

GPUs for AI in 2026: NVIDIA, AMD, Intel Compared

NVIDIA, AMD, and Intel compete in the 2026 AI GPU market, with NVIDIA's Blackwell RTX 50-series, AMD's Radeon AI Pro R9700, and Intel's Arc Pro B70 targeting local LLM inference. VRAM capacity and mem…

← prev page 10 / 16 next →
// co-occurs with top 8 entities
// topics top 6 topics