MiniMax

MiniMax M2.5 — Benchmark Scores, Pricing & Performance Analysis

FLAGSHIPMiniMax
Chatbot Arena ELO
1443
Output Speed
59 tok/s
Input Cost
$0.27/1M
Output Cost
$1.1/1M
Context Window
205K
Max Output
131K
Accepts
text

MiniMax M2.5 by MiniMax demonstrates top-tier general intelligence, competitive pricing. View detailed benchmark data including scores across coding, math, reasoning, speed, and cost metrics.

Specifications last checked against the provider on . Benchmark scores carry their own source, linked beside each number.

General Benchmarks

Chatbot Arena ELO
144312th of 74

Human preference ELO from blind head-to-head votes

LMSYS / HuggingFace

Coding Benchmarks

Design Arena ELO
120044th of 120

Human preference ELO for building things — websites, UI, games, charts, SVG

Design Arena (via OpenRouter)
Website Arena ELO
123441st of 108

Human preference ELO for building a working web page from a brief

Design Arena (via OpenRouter)

Reasoning Benchmarks

τ²-Bench Airline
64.1%63rd of 93

Multi-turn service agent making tool calls under strict policy constraints

OpenRouter (measured)
GPQA Diamond
84.1%40th of 140

Graduate-level science Q&A by domain experts

Papers

Speed Benchmarks

Output Speed
59 tok/s72nd of 94

Tokens generated per second

Aggregated
Time to First Token
2330 ms88th of 92

Latency before first token arrives

Aggregated

Cost Benchmarks

Cached Input Cost
$0.0324th of 136

Cost per 1M cached input tokens — the price that actually applies to a long agent conversation

OpenRouter API
Output Cost
$1.1205th of 413

Cost per 1M output tokens

OpenRouter API
Input Cost
$0.27202nd of 413

Cost per 1M input tokens

OpenRouter API

Context Benchmarks

Context Length
205K177th of 451

Maximum context window size

OpenRouter API

Available from 45 providers

The same model costs different amounts depending on who serves it. You can bring your own key for 8 of these — connect it here.

ProviderInput / 1MOutput / 1MCached inContextYour key
Alibaba Coding Plan$0$0$0197K—
Alibaba Coding Plan (China)$0$0$0197K—
Alibaba Token Plan$0$0$0197K—
Alibaba Token Plan (China)$0$0$0197K—
MiniMax Token Plan (minimaxi.com)$0$0$0205K—
MiniMax Token Plan (minimax.io)$0$0$0205K—
SCNet Token Plan$0$0$0205K—
Tencent Coding Plan (China)$0$0$0205K—
Deep Infra$0.15$1.15$0.03197KSupported
GreenPT$0.1938$1.129$0.0627205K—
DInference$0.22$0.88—200K—
OpenRouter$0.27$1.08$0.027205KSupported
Venice AI$0.27$0.95$0.03198K—
D.Run (China)$0.29$1.16—205K—
Cortecs$0.296$1.186$0.075196K—
302.AI$0.3$1.2—205K—
Alibaba (China)$0.3$1.2—205K—
Amazon Bedrock$0.3$1.2—197KSupported
Baseten$0.3$1.2—204KSupported
CloudFerro Sherlock$0.3$1.2—196K—
DevPass (LLM Gateway)$0.3$1.2$0.03229K—
DigitalOcean$0.3$1.2$0.0666K—
Eden AI$0.3$1.2$0.03205K—
Friendli$0.3$1.2$0.06197K—
FrogBot$0.3$1.2$0.03192K—
HPC-AI$0.3$1.2$0.03196K—
Hugging Face$0.3$1.2$0.03205KSupported
Kilo Gateway$0.3$1.2$0.03200K—
LLM Gateway$0.3$1.2$0.03205K—
Meganova$0.3$1.2—205K—
Merge Gateway$0.3$1.2$0.03205K—
MiniMax (minimaxi.com)$0.3$1.2$0.03205K—
MiniMax (minimax.io)$0.3$1.2$0.03205KSupported
NanoGPT$0.3$1.2$0.15205K—
NovitaAI$0.3$1.2$0.03205K—
Ofox$0.3$1.2$0.03205K—
OpenCode Go$0.3$1.2$0.03205K—
OpenCode Zen$0.3$1.2$0.06205K—
OrcaRouter$0.3$1.2$0.03205K—
TensorX$0.3$1.2$0.075197K—
Together AI$0.3$1.2$0.06205KSupported
TokenGo$0.3$1.2$0.03205K—
Vercel AI Gateway$0.3$1.2$0.03205KSupported
ZenMux$0.3$1.2$0.03205K—
Ollama Cloud———205K—

MiniMax M2.5 — Benchmark Scores Overview

Scores normalized to percentage scale for visual comparison. ELO scores mapped to 0-100 range (1100-1500).

MiniMax M2.5 — Frequently Asked Questions

How intelligent is MiniMax M2.5?

MiniMax M2.5 scores 1443 on the Chatbot Arena ELO rating, making it a high-performing AI model. This score is based on blind head-to-head human preference voting.

How much does MiniMax M2.5 cost?

MiniMax M2.5 costs $0.27 per 1M input tokens and $1.1 per 1M output tokens. This makes it one of the more affordable models.

How fast is MiniMax M2.5?

MiniMax M2.5 generates output at 59 tokens per second, which is slower, prioritizing quality over speed compared to other models. The time to first token is 2330 ms.

What is the context window of MiniMax M2.5?

MiniMax M2.5 has a context window of 205K tokens. This determines how much text, conversation history, and code the model can process in a single request.

Who created MiniMax M2.5?

MiniMax M2.5 was created by MiniMax. It is classified as a flagship model in our catalogue.

Is MiniMax M2.5 open source?

No, MiniMax M2.5 is a proprietary model. It is available through MiniMax's API and compatible providers.