cd/entity/MiniMax-M3· home entities MiniMax-M3
grep -l @minimax-m3 /news/*.json | wc -l → 18

MiniMax-M3

mentions 18 type Organization feed RSS

// recent coverage 18 mentions

02:18
2026-07-24
workbuddybench.com
artificial-intelligence

Tencent WorkBuddy Bench – Agentic Coding Leaderboard

Tencent's WorkBuddy Bench leaderboard shows no single model dominates agentic coding tasks, with Claude Opus 4.8 leading five of eight scored columns, GLM-5.2 leading two, and GPT-5.5 leading one. The…

19:54
2026-07-14
machinebrief.com
artificial-intelligence

StructAgent: Revolutionizing Long-Horizon Task Management

StructAgent, a new state-centered framework for long-horizon task management, achieved a 78.9% success rate on the OSWorld-Verified platform using the MiniMax-M3 model, setting a new open-source state…

06:09
2026-07-14
artificialanalysis.ai
artificial-intelligence

Harvey LAB-AA: evaluating AI agents on real-world legal work

Harvey LAB-AA, a new benchmark from Artificial Analysis evaluating AI agents on real-world legal work across 24 practice areas, shows Claude Fable 5 (max, with Opus 4.8 fallback) leading with a 14.2% …

14:12
2026-07-09
tokenstead.ai
artificial-intelligence

Open-weights models cost less: a 2026 pricing guide

Open-weights models are dramatically cheaper per token than closed models, with the five cheapest models on the Artificial Analysis pricing index all being open-weights and the five most expensive all…

12:00
2026-06-23
telnyx.com
large-language-models

GLM-5.2 is now available on Telnyx Inference

Telnyx has added Z.ai's GLM-5.2 open-weight model to its Inference platform, hosted on owned B300 GPUs. The model leads Artificial Analysis with an Intelligence Index of 51, outperforming competitors,…

00:00
2026-06-22
runagentrun.co.uk
ai-agents

A business assistant for under £50 a month

A new open-source AI assistant stack combining the Hermes agent and MiniMax-M3 model on Nous Portal costs under £50 per month and can automate tasks like market briefings, inbox triage, lead research,…

00:00
2026-06-21
epics.tech
large-language-models

Agents That Fix Themselves, and the Collapse of the Scale Law

On June 21, GLM-5.2, an open-weight model with 40 billion active parameters, outperformed larger closed models on the Intelligence Index, signaling a collapse of the scale law. Concurrently, research …

23:57
2026-06-18
artificialanalysis.ai
ai-research

Show HN: AA-Briefcase: a frontier knowledge work evaluation

A new evaluation benchmark, AA-Briefcase, measures frontier knowledge work performance, with models like Claude Opus 4.8 averaging 24 minutes per task and achieving an Elo of 1356, while MiniMax-M3 ta…

23:58
2026-06-17
simonwillison.net
large-language-models

GLM-5.2 is probably the most powerful text-only open weights LLM

Chinese AI lab Z.ai released GLM-5.2, a 753B-parameter open-weights text-only LLM with a 1 million token context window, under an MIT license. The model leads the Artificial Analysis Intelligence Inde…

12:00
2026-06-12
arxiv.org
machine-learning

Maxproof

Researchers have developed MaxProof, a population-level test-time scaling framework for mathematical proof that enables the MiniMax-M3 model to achieve 35 out of 42 on IMO 2025 and 36 out of 42 on USA…

// co-occurs with top 8 entities
// topics top 6 topics