cd/entity/CUDA· home entities CUDA
grep -l @cuda /news/*.json | wc -l → 319

CUDA

mentions 319 type Organization page 12/16 feed RSS

// recent coverage 319 mentions

14:09
2026-06-28
byteiota.com
artificial-intelligence

Qualcomm’s $3.9B Modular Buy: Mojo Takes Aim at CUDA

Qualcomm acquired AI startup Modular for $3.9 billion on June 24, 2026, to challenge NVIDIA's CUDA software dominance with Modular's Mojo programming language and MAX inference engine, which enable AI…

04:08
2026-06-27
dev.to
developer-tools

Headless Mode on NVIDIA Jetson AGX Orin 64GB with JetPack 7.2

A developer documented a method to enable fully headless operation on the NVIDIA Jetson AGX Orin 64GB Developer Kit running JetPack 7.2. The guide uses NoMachine for remote GUI access with XFCE4 inste…

07:02
2026-06-26
dev.to
large-language-models

Self-Hosted Ollama Homelab: 3 Mistakes Running Local LLMs

A developer's team encountered three critical mistakes while setting up a self-hosted Ollama homelab for DevOps assistance: silent CPU fallback due to CUDA version mismatch, an OOM crash that took dow…

06:22
2026-06-25
koreatimes.co.kr
artificial-intelligence

Why did Jensen Huang refuse to see the future?

Nvidia CEO Jensen Huang told tvN's 'You Quiz on the Block' that he would choose resilience over perfect foresight, and speaking with his future self over his past self. His decisions, such as investin…

05:22
2026-06-25
dev.to
artificial-intelligence

How to Transcribe Meetings Locally in 2026 (Whisper, On-Device)

Off Grid AI Desktop is a free, open-source app that records and transcribes meetings locally on Mac or PC using OpenAI's Whisper model, ensuring audio never leaves the machine and eliminating per-minu…

05:12
2026-06-25
byteiota.com
artificial-intelligence

Qualcomm Acquires Modular: Mojo, MAX, and CUDA’s Future

Qualcomm acquired Modular for $3.9 billion in stock, gaining its Mojo programming language and MAX inference platform to challenge NVIDIA's CUDA dominance. The deal, expected to close in the second ha…

20:50
2026-06-24
discuss.huggingface.co
large-language-models

Niodoo hidden state steering

Developer Jason Van Pham released Niodoo, a runtime that uses hidden state steering to improve small language models' performance without fine-tuning, enabling self-correction and memory systems. The …

15:54
2026-06-24
polarsignals.com
developer-tools

Optimizing a CUDA FSST decompression kernel

Polar Signals storage engineer wrote a CUDA kernel to decompress FSST-encoded strings on GPUs, but initial performance was worse than CPU decompression. Using the company's GPU profiler with PC sampli…

14:52
2026-06-24
sipp.sh
large-language-models

Show HN: Sipp – Run small local LLMs in browser 3x faster

Sipp, an open-source AI inference library, enables running small local LLMs in browsers with up to 3x faster decode speeds than alternatives. Built by HCI and graphics programmers, it uses a unified A…

12:53
2026-06-24
cast.ai
artificial-intelligence

TPUs vs GPUs: When to Choose What for AI/ML Workloads

Google's TPU v5p and upcoming Ironwood TPU7x offer higher raw matrix throughput and simpler interconnects than NVIDIA H100 GPUs for large-scale AI training, but GPUs remain superior for inference with…

10:00
2026-06-24
cio.com
ai-infrastructure

Choosing your AI stack: The benefits of vendor lock-in

Accenture research shows 97% of executives believe AI will transform their companies, but CIOs are discovering that AI stack decisions are not easily reversible due to tight co-engineering across comp…

← prev page 12 / 16 next →
// co-occurs with top 8 entities
// topics top 6 topics