cd/entity/CUDA· home entities CUDA
grep -l @cuda /news/*.json | wc -l → 318

CUDA

mentions 318 type Organization page 1/16 feed RSS

// recent coverage 318 mentions

02:09
2026-09-11
byteiota.com
ai-products

Mojo 1.0 Is Fully Open Source: What Developers Must Do Now

Qualcomm-owned Modular released Mojo 1.0 on August 11 with its first stability guarantees and open-sourced the compiler, toolchain, and standard library under Apache 2.0 with LLVM exceptions on August…

20:30
2026-09-10
dev.to
robotics

Building a Vision-Language Robot with Jetson + ROS 2

A developer published a tutorial on building a vision-language robot pipeline that combines camera perception with natural-language instructions on NVIDIA Jetson hardware running ROS 2. The guide walk…

18:54
2026-09-10
frontierroles.com
ai-infrastructure

Inference Engineering, Co-op — Inferact

Inferact, founded by the creators and core maintainers of vLLM, is hiring University of Waterloo co-op students for an on-site Inference Engineering co-op in San Francisco, with pay not published. The…

22:09
2026-09-08
frontierroles.com
artificial-intelligence

Staff Applied AI Inference Engineer — Crusoe

Crusoe, a vertically integrated AI infrastructure company, is hiring a Staff Applied AI Inference Engineer in San Francisco with a salary range of $215,000–260,000 per year, which sits 8% above the $2…

05:31
2026-09-08
leimao.github.io
ai-infrastructure

CUDA Multi-Process Service

NVIDIA's CUDA Multi-Process Service (MPS) enables concurrent kernel execution across multiple processes, improving GPU utilization compared to default time-slicing, according to a technical blog post …

10:45
2026-09-07
discuss.huggingface.co
artificial-intelligence

RuntimeError: No CUDA GPUs are available and ValueError

A user named Stormpie reported on a forum that their pottery dating machine 'terraID' fails with 'RuntimeError: No CUDA GPUs are available' when clicking the 'Analyze' button, despite the app running …

09:00
2026-09-07
github.com
ai-infrastructure

AI performance engineering: a primary-source reading list

Wafer, an AI company, published a curated reading list for learning GPU performance engineering and production inference, ordered from a single inference request to distributed systems and featuring p…

07:39
2026-09-04
promptcube3.com
ai-infrastructure

Optimizing CUDA kernels manually is becoming a specialized art

Optimizing CUDA kernels manually is becoming a specialized art, with developers moving beyond basic implementations to structured workflows that prioritize profiling-first approaches using NVIDIA Nsig…

06:31
2026-09-04
github.com
machine-learning

Nano LM Studio

NanoLM Studio V4, a local desktop workbench for building small decoder-only language models, has been publicly released by its developer. The Tk-based application integrates document ingestion, ByteLe…

00:06
2026-09-04
promptcube3.com
ai-infrastructure

Nvidia might actually buy Hugging Face to dominate the AI stack

Nvidia may acquire Hugging Face to create a unified AI stack from silicon to deployed models, according to an analysis. The potential merger would streamline model optimization, compute integration, a…

17:19
2026-09-03
dev.to
ai-infrastructure

Deploying Inference Using NVIDIA Dynamo and vLLM

NVIDIA Dynamo, an open-source inference framework, has been deployed with the vLLM backend to serve chat completion requests in both aggregated and disaggregated configurations. The deployment involve…

08:19
2026-09-03
github.com
large-language-models

Show HN: PicoLM v1.0-rc1

PicoLM v1.0-rc1, an LLM inference engine written in C99, has been released, supporting llama-2, GPT-2, Qwen 3.6/3.8(+MoE), and Gemma-3n models. The engine features CPU SIMD acceleration, CUDA/HIP supp…

12:30
2026-09-02
iaroslavelistratov.github.io
artificial-intelligence

B200 Attention Kernel from Scratch to Near-SOTA in 60 Diagrams

A new technical blog post by Iaroslav Elistratov presents a visual guide to building a B200 attention kernel from scratch in CUDA and PTX, achieving 94.4% of FlashAttention-4 performance on 4K, 8K, an…

19:00
2026-09-01
servethehome.com
artificial-intelligence

NVIDIA RISC-V for NVIDIA GPUs at Hot Chips 2026

At Hot Chips 2026, NVIDIA detailed its plans to bring RISC-V CPUs into its GPU platforms, outlining requirements for CUDA and NVLink Fusion integration. The company highlighted RVA23 profiles and the …

16:32
2026-09-01
eetimes.com
artificial-intelligence

How AI Is Reshaping the Global Semiconductor Patent Landscape

Patent filings at the intersection of AI and semiconductors increased 114% over the past five years, outpacing the broader sector's 78% growth, according to a report by IP management provider Anaqua. …

page 1 / 16 next →
// co-occurs with top 8 entities
// topics top 6 topics