cd/entity/Mixture of Experts· home entities Mixture of Experts
grep -l @mixture of experts /news/*.json | wc -l → 8

Mixture of Experts

mentions 8 type Person feed RSS

// recent coverage 8 mentions

12:50
2026-09-15
shiftmag.dev
large-language-models

LLMs Have a New Limit – It Costs More to Think Longer

Nvidia Developer Relations Manager Igor Dmochowski said at Infobip Shift 2026 that the compute required to process an LLM's context grows quadratically with context length, making it unscalable for pr…

19:12
2026-09-13
arxiv.org
large-language-models

MOBA: Mixture of Block Attention for Long-Context LLMs

MoonshotAI researchers submitted a paper on 18 Feb 2025 introducing Mixture of Block Attention (MoBA), an attention mechanism that applies Mixture of Experts principles to let long-context large langu…

23:49
2026-06-16
arxiv.org
large-language-models

The Guide to Fine-Tuning LLMs

A comprehensive review published on arXiv examines fine-tuning techniques for Large Language Models (LLMs), covering methodologies from supervised and unsupervised learning to parameter-efficient meth…

// co-occurs with top 8 entities
// topics top 6 topics