cd/entity/CUDA· home entities CUDA
grep -l @cuda /news/*.json | wc -l → 318

CUDA

mentions 318 type Organization page 4/16 feed RSS

// recent coverage 318 mentions

12:19
2026-08-19
yassa9.github.io
computer-vision

Geolocating a random island using geometry and CUDA programming

A developer known as gralhix solved an OSINT geolocation challenge by writing CUDA-accelerated geometry code instead of using Google Lens, identifying a resort island from a drone photo. The solution,…

21:04
2026-08-18
promptcube3.com
ai-infrastructure

PantheonGPU proves that telemetry alone is a lie for GPU health

PantheonGPU, a new GPU health-check tool, runs 45+ targeted tests on NVIDIA CUDA and AMD ROCm hardware to detect stability issues and configuration bottlenecks that telemetry tools miss. The tool stre…

21:01
2026-08-18
pub.towardsai.net
artificial-intelligence

What a Kernel Is, and Why Everyone Is Writing New Ones

A kernel is a single function that runs on a GPU, and the steep memory hierarchy—with a 1,600x gap between L2 cache and HBM—makes kernel design crucial for AI performance. FlashAttention, developed by…

16:33
2026-08-18
promptcube3.com
ai-infrastructure

Groq is spending billions to poach Nvidia engineers

Groq is spending billions to poach Nvidia engineers as it targets the AI inference market, aiming to challenge Nvidia's dominance by optimizing hardware for the sequential nature of LLM token generati…

23:11
2026-08-17
byteiota.com
artificial-intelligence

Rust GPU Offload Hits rustc: Safe, Portable Kernels Now

A research paper, “GPU Offload in Rust: Portable, Safe, and Fast” (arXiv:2608.13759), introduces GPU programming support directly into rustc, allowing developers to write GPU kernels in safe Rust with…

23:11
2026-08-17
jonidimo.github.io
large-language-models

Qwen3.8-27B on a single RTX 3090: crash fix, 131K context, 9 myths

A 14-hour benchmark on a single NVIDIA GeForce RTX 3090 found that the Qwen3.8-27B hybrid SSM+attention model achieves a 131K context window on 24 GB VRAM, scoring 20/21 on a frontier test set, but cr…

10:14
2026-08-17
byteiota.com
artificial-intelligence

Mojo 1.0 Is Here: Python Speed, Rust Safety, AI Hardware

Modular shipped Mojo 1.0 on August 12 as part of its 26.5 release, marking the first production-ready version of its Python-syntax AI hardware language, with changes including uniform `var` syntax, a …

03:27
2026-08-16
promptcube3.com
artificial-intelligence

Moving 250k lines of legacy weather simulation code to GPUs

A developer reports that porting 250,000 lines of legacy weather simulation code to GPUs requires a chunk-and-verify workflow using an LLM agent to map data flow, identify hot kernels, and iteratively…

12:31
2026-08-15
promptcube3.com
artificial-intelligence

What would you actually build if you had a stack of GPUs and

A developer argues that high-VRAM GPUs like Nvidia H100s or RTX 4090s should be used for compute-heavy tasks beyond local LLMs, such as real-time fluid dynamics, custom GAN training on niche datasets,…

05:16
2026-08-15
promptcube3.com
ai-infrastructure

Nvidia is chasing a 500 billion dollar target that has Wall

Nvidia is pursuing a $500 billion market valuation target by integrating its InfiniBand networking, CUDA software, and Blackwell architecture into a proprietary full-stack ecosystem, creating high swi…

17:32
2026-08-14
promptcube3.com
large-language-models

Why is Qwen 2.5-27B acting so erratic on my local setup?

A developer reports that Qwen 2.5-27B, a large language model, exhibits erratic behavior and degraded output quality on a local setup with an NVIDIA RTX 3090 (24GB VRAM), particularly when the context…

21:32
2026-08-13
promptcube3.com
artificial-intelligence

Reproducing 2

A new analysis of 2,200 AI research reproducibility cases finds that most papers fail to replicate due to hardware variance, hyper-parameter sensitivity, and dependency issues, with the missing link o…

← prev page 4 / 16 next →
// co-occurs with top 8 entities
// topics top 6 topics