cd/entity/RCCL· home› entities› RCCL
grep -l @rccl /news/*.json | wc -l → 8

RCCL

mentions 8 type Organization feed RSS

// recent coverage 8 mentions

09:29
2026-08-26
forum.level1techs.com
large-language-models

24/32GB GPU in an AM4 system?

A forum user argues that 32GB of VRAM is the practical threshold for useful large language model use, noting that 24GB cards work but with limited context size. The user highlights that NVIDIA offers …

19:57
2026-08-20
forum.level1techs.com
artificial-intelligence

DeepSeek V4 Flash on 8× AMD gfx1201: packaged TP=8 deployment

DeepSeek V4 Flash, a 284B-parameter mixture-of-experts model with 256 routed experts and FP4 expert weights, was successfully deployed on eight AMD Radeon AI PRO R9600D GPUs (32 GB each, 256 GB total)…

00:00
2026-07-13
rocm.blogs.amd.com
artificial-intelligence

QuickReduce INT3 Quantization and Benchmarking on MI355

AMD's QuickReduce library now supports INT3 quantization for all-reduce communication in multi-GPU LLM inference, achieving a 22% reduction in on-wire data volume compared to INT4 on AMD Instinct MI35…

00:46
2026-06-28
github.com
artificial-intelligence

AMD Strix Halo RDMA Cluster Setup Guide

AMD Strix Halo cluster setup guide details how to configure a two-node system linked via Intel E810 RoCE v2 for distributed vLLM inference using Tensor Parallelism. The guide covers hardware prerequis…

03:53
2026-06-16
blog.maincode.com
artificial-intelligence

Writing our own inference engine in Rust on the AMD MI355X

Maincode, an Australian AI company, built a custom inference engine in Rust for the AMD MI355X GPU that bypasses the ROCm runtime entirely, achieving lower overhead on critical operations like tensor-…

// co-occurs with top 8 entities
// topics top 6 topics