cd/entity/MiniMax-M3· home entities MiniMax-M3
grep -l @minimax-m3 /news/*.json | wc -l → 24

MiniMax-M3

mentions 24 type Organization page 1/2 feed RSS

// recent coverage 24 mentions

12:47
2026-09-03
alphaxiv.org
artificial-intelligence

Harness-of-Harness: Multi-Day Autonomous Software Development

Researchers from Shanghai Artificial Intelligence Laboratory and National University of Singapore introduced Harness-of-Harness (HoH), a framework that enables LLM-based coding agents to autonomously …

08:36
2026-08-28
artificialanalysis.ai
artificial-intelligence

Agnes AI Releases Agnes 2.5 Pro Beta

Agnes AI's Agnes 2.5 Pro Beta scores 49 on the Artificial Analysis Intelligence Index, up 9 points from Agnes 2.5 Pro Alpha, driven by large agentic gains but using ~2x the output tokens. The beta che…

11:54
2026-08-01
runinfra.ai
ai-infrastructure

$0.09 and $290.12 are both the price of 1M output tokens

A cost analysis across 24 providers and 378 GPU rental rates found that the price of one million output tokens ranges from $0.09 on a single AMD MI355X to $290.12 on eight NVIDIA H100s, with the gap d…

02:18
2026-07-24
workbuddybench.com
artificial-intelligence

Tencent WorkBuddy Bench – Agentic Coding Leaderboard

Tencent's WorkBuddy Bench leaderboard shows no single model dominates agentic coding tasks, with Claude Opus 4.8 leading five of eight scored columns, GLM-5.2 leading two, and GPT-5.5 leading one. The…

19:54
2026-07-14
machinebrief.com
artificial-intelligence

StructAgent: Revolutionizing Long-Horizon Task Management

StructAgent, a new state-centered framework for long-horizon task management, achieved a 78.9% success rate on the OSWorld-Verified platform using the MiniMax-M3 model, setting a new open-source state…

06:09
2026-07-14
artificialanalysis.ai
artificial-intelligence

Harvey LAB-AA: evaluating AI agents on real-world legal work

Harvey LAB-AA, a new benchmark from Artificial Analysis evaluating AI agents on real-world legal work across 24 practice areas, shows Claude Fable 5 (max, with Opus 4.8 fallback) leading with a 14.2% …

14:12
2026-07-09
tokenstead.ai
artificial-intelligence

Open-weights models cost less: a 2026 pricing guide

Open-weights models are dramatically cheaper per token than closed models, with the five cheapest models on the Artificial Analysis pricing index all being open-weights and the five most expensive all…

12:00
2026-06-23
telnyx.com
large-language-models

GLM-5.2 is now available on Telnyx Inference

Telnyx has added Z.ai's GLM-5.2 open-weight model to its Inference platform, hosted on owned B300 GPUs. The model leads Artificial Analysis with an Intelligence Index of 51, outperforming competitors,…

00:00
2026-06-22
runagentrun.co.uk
ai-agents

A business assistant for under £50 a month

A new open-source AI assistant stack combining the Hermes agent and MiniMax-M3 model on Nous Portal costs under £50 per month and can automate tasks like market briefings, inbox triage, lead research,…

00:00
2026-06-21
epics.tech
large-language-models

Agents That Fix Themselves, and the Collapse of the Scale Law

On June 21, GLM-5.2, an open-weight model with 40 billion active parameters, outperformed larger closed models on the Intelligence Index, signaling a collapse of the scale law. Concurrently, research …

23:57
2026-06-18
artificialanalysis.ai
ai-research

Show HN: AA-Briefcase: a frontier knowledge work evaluation

A new evaluation benchmark, AA-Briefcase, measures frontier knowledge work performance, with models like Claude Opus 4.8 averaging 24 minutes per task and achieving an Elo of 1356, while MiniMax-M3 ta…

23:58
2026-06-17
simonwillison.net
large-language-models

GLM-5.2 is probably the most powerful text-only open weights LLM

Chinese AI lab Z.ai released GLM-5.2, a 753B-parameter open-weights text-only LLM with a 1 million token context window, under an MIT license. The model leads the Artificial Analysis Intelligence Inde…

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics