cd/entity/Mixture-of-Experts· home entities Mixture-of-Experts
grep -l @mixture-of-experts /news/*.json | wc -l → 16

Mixture-of-Experts

mentions 16 type Organization feed RSS

// recent coverage 16 mentions

09:13
2026-08-11
vincentschmalbach.com
large-language-models

How Does Mixture-of-Experts Routing Affect LLM Repeatability?

A new analysis from the Journal of Machine Learning Research explains that Mixture-of-Experts (MoE) routing can cause large language models to produce different outputs across runs due to discrete exp…

04:00
2026-08-03
arxiv.org
artificial-intelligence

Topology-Aware Data Movement for Disaggregated GPU Inference

A new arXiv paper (arXiv:2607.28633v1) proposes a topology-aware transfer orchestrator for disaggregated GPU inference, claiming existing systems like DistServe, Splitwise, and Mooncake ignore that ba…

20:59
2026-07-27
cryptobriefing.com
artificial-intelligence

Kimi K3 leads open-weight models in Agent Arena with +10% score

Moonshot AI's Kimi K3 (Max), a 2.8-trillion-parameter open-weight model, achieved a +9.75% net-improvement score to rank third overall in Agent Arena, trailing only Claude Fable 5 (High) and GPT-5.6 S…

13:56
2026-07-21
arxiv.org
artificial-intelligence

Loop the Loopies

Researchers have introduced the Loopie series, two Mixture-of-Experts models (20B parameters with 2B active and 6B with 0.6B active) that outperform vanilla Transformer baselines trained with the same…

04:00
2026-07-07
arxiv.org
large-language-models

Gemma 4 Technical Report

Google DeepMind released Gemma 4, a new generation of open-weight multimodal language models featuring dense and Mixture-of-Experts architectures from 2.3B to 31B parameters. The models introduce a th…

16:00
2026-06-24
huggingface.co
large-language-models

Accelerating Transformers Fine-Tuning with NVIDIA NeMo AutoModel

NVIDIA released NeMo AutoModel, an open-source library that accelerates fine-tuning of Mixture-of-Experts (MoE) transformer models by 3.4-3.7x in training throughput and reduces GPU memory usage by 29…

04:00
2026-06-15
arxiv.org
large-language-models

Can Editing 1 Neuron Fix Repetition Loops in LLMs?

Researchers found that repetition loops in Gemma 4 instruction-tuned LLMs can be fixed by editing as few as one neuron, but the fix does not resolve deeper 'doom loops' caused by missing knowledge. Th…

// co-occurs with top 8 entities
// topics top 6 topics