cd/entity/CUDA· home entities CUDA
grep -l @cuda /news/*.json | wc -l → 318

CUDA

mentions 318 type Organization page 3/16 feed RSS

// recent coverage 318 mentions

14:33
2026-08-25
metalworking.vercel.app
machine-learning

GPU Glossary but for Apple Silicon GPUs

Apple's Metal GPU glossary 'metalworking' targets developers familiar with CUDA but new to Apple Silicon, mapping every concept from GPU cores to MLX's architecture with CUDA equivalents and real kern…

09:29
2026-08-25
lightreading.com
artificial-intelligence

Ericsson, Nokia and Samsung clash over 6G's need for Nvidia

Nokia's chief technology officer Pallavi Mahajan said the company's $1 billion investment from Nvidia will enable GPU-based radio access network (RAN) software that can run compute-intensive machine-l…

20:53
2026-08-24
promptcube3.com
artificial-intelligence

I'll skip the clickbait headline rewrite here because this

Nvidia's Jetson Orin edge AI module, delivering up to 275 TOPS, is widely used in robotics and embedded systems, including drones and computer vision applications, according to Nvidia's official speci…

20:12
2026-08-24
byteiota.com
artificial-intelligence

NVIDIA CUDA Targets RISC-V: What the Server Play Means for Devs

Nvidia announced at Hot Chips 2026 that CUDA will officially support RISC-V CPUs, making RISC-V the third supported host architecture alongside x86 and ARM, with strict requirements including the RVA2…

16:52
2026-08-24
chipsandcheese.com
artificial-intelligence

Hot Chips 2026: CUDA Targets RISC-V – By Chester Lam

Nvidia is extending CUDA support to RISC-V CPUs, requiring server-grade features such as RVA23 compliance, ACPI, PCIe coherency, and peer-to-peer PCIe, according to a Hot Chips 2026 talk. The company …

10:30
2026-08-24
dev.to
artificial-intelligence

jina-embeddings-v4 as an OpenAI-Compatible Embeddings Server

A developer has released jina-embeddings-v4, a self-hosted server for the jina-embeddings-v4 embedding model with an OpenAI-compatible /v1/embeddings endpoint. The server runs on a single NVIDIA GPU a…

07:00
2026-08-24
hiraditya.github.io
artificial-intelligence

A Bug Is a Violation of a Specification

A bug is a violation of a specification, and no specification exists that prefix caching's variable logits violate, according to an analysis of vLLM and SGLang issues. The vLLM PR #34046 adds an opt-i…

06:22
2026-08-24
frontierroles.com
machine-learning

ML Systems Performance Engineer (MFU) — Higgsfield

Higgsfield AI, a generative AI company with $500M in annual revenue run rate and 25M+ users, is hiring an ML Systems Performance Engineer (MFU) for its Almaty, Kazakhstan office. The role focuses on o…

23:56
2026-08-23
promptcube3.com
machine-learning

Picking a major for Physics-Informed Neural Networks (PINNs) is

A technical essay argues that mastering Physics-Informed Neural Networks (PINNs) requires a hybrid background rather than a single major, recommending Computer Science or Mathematics paired with elect…

22:00
2026-08-23
fratepietro.com
artificial-intelligence

Ferrox v0.9.1: a Rust GGUF engine, measured against llama.cpp

Ferrox v0.9.1, a pure-Rust GGUF inference engine, shipped today with MoE prefill on Apple Metal 2.4x faster, closing the gap to llama.cpp from 2.62x behind to 1.11x on OLMoE-1B-7B. The update also imp…

05:55
2026-08-21
github.com
developer-tools

Testing CUDA kernel execution with GPU correctness harness

A new CUDA kernel correctness harness compiles kernels with nvcc, loads them into Python via ctypes, and checks outputs against NumPy baselines, with timing via CUDA events. All 23 tests pass on an NV…

19:02
2026-08-19
promptcube3.com
artificial-intelligence

Cerebras WSE-3 smokes H100 on Llama 3 70B inference at a

Cerebras Systems claims its Wafer-Scale Engine 3 (WSE-3) delivers 1,800 tokens per second at 23 kW for Llama 3 70B inference, versus about 850 tokens per second at 56 kW for an 8×H100 DGX system, yiel…

18:38
2026-08-19
dev.to
developer-tools

Finding a Random Island with Geometry and CUDA

A developer detailed a CUDA-based approach to finding the closest island to any ocean point using spherical geometry and GPU acceleration. The method leverages the Haversine formula and spatial indexi…

15:35
2026-08-19
gist.github.com
developer-tools

Hashcat on DGX Spark (GB10) - MSI EdgeExpert

A developer benchmarked Hashcat on an NVIDIA GB10 GPU inside a DGX Spark system, reporting speeds of 81.16 GH/s for MD4, 53.54 GH/s for MD5, and 16.70 GH/s for SHA1. The benchmark used Hashcat v7.1.2-…

← prev page 3 / 16 next →
// co-occurs with top 8 entities
// topics top 6 topics