cd/entity/NVIDIA· home entities NVIDIA
grep -l @nvidia /news/*.json | wc -l → 2069

NVIDIA

mentions 2069 type Organization page 52/104 feed RSS
sameAs · en.wikipedia.org · www.wikidata.org

// recent coverage 2069 mentions

00:00
2026-07-13
rocm.blogs.amd.com
artificial-intelligence

Serving NVFP4 Models on AMD Instinct™ MI355 Accelerators

AMD has integrated an NVFP4 emulation pipeline into vLLM that enables AMD Instinct MI355 accelerators to serve standard NVFP4 quantized checkpoints directly, dequantizing weights to BF16 on-the-fly at…

00:00
2026-07-13
runagentrun.co.uk
ai-infrastructure

NVIDIA Vera targets the agent-loop bottleneck

NVIDIA published a new CPU category on 7 July with Vera, an Arm server CPU designed to address the agentic-AI bottleneck by maximizing single-threaded performance rather than core density. In testing …

23:31
2026-07-12
pub.towardsai.net
artificial-intelligence

AI’s Biggest Bottleneck Isn’t GPUs Anymore

SK Hynix raised $26.5 billion in the largest foreign IPO in US history, signaling investor demand for memory chips over GPUs. NVIDIA's stock fell 15% while memory maker Micron tripled, and H100 GPU re…

12:22
2026-07-12
blawg.pages.dev
artificial-intelligence

Ollama vs. Llama.cpp – Quick Benchmark

A benchmark comparing Ollama and llama-server on a Tesla V100 found that both backends achieve nearly identical token generation speeds (about 111 tok/s), but llama-server processed prompts 22% faster…

05:51
2026-07-12
gist.github.com
artificial-intelligence

vllm locally on 5060Ti 16GB x 2

A developer deployed vLLM locally on two NVIDIA RTX 5060 Ti 16GB GPUs using Docker Compose, configuring tensor parallelism, FP8 KV cache, and speculative decoding with MTP. The setup runs an OpenAI-co…

00:04
2026-07-12
sourcefeed.dev
large-language-models

Fine-Tune Qwen2.5-7B with QLoRA on Your Own Data

Mariana Souza published a practical guide for fine-tuning Qwen2.5-7B-Instruct using QLoRA on custom instruction datasets, including cost estimates and a loss-masking sanity check. The tutorial covers …

22:04
2026-07-11
sourcefeed.dev
ai-infrastructure

Demystifying the NVIDIA DGX Spark for API Developers

NVIDIA's DGX Spark desktop GPU, with 128 GB unified memory and a 140W ARM64 processor, challenges API developers to shift from cloud-based AI consumption to local systems engineering. The device's sha…

12:52
2026-07-11
machinebrief.com
robotics

Why Virtual Gyms Are Shaping the Future of Robotics

Virtual gyms are high-fidelity simulation environments that train robots using digital twins, synthetic data, and reinforcement learning, offering a safer and more cost-effective alternative to real-w…

← prev page 52 / 104 next →
// co-occurs with top 8 entities
// topics top 6 topics