Back to Models

MiniMax: MiniMax M2.1Active

minimax/minimax-m2.1
CompareChat

MiniMax-M2.1 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application development. With only 10 billion activated parameters, it delivers a major jump in real-world capability while maintaining exceptional latency, scalability, and cost efficiency.

Compared to its predecessor, M2.1 delivers cleaner, more concise outputs and faster perceived response times. It shows leading multilingual coding performance across major systems and application languages, achieving 49.4% on Multi-SWE-Bench and 72.5% on SWE-Bench Multilingual, and serves as a versatile agent “brain” for IDEs, coding tools, and general-purpose assistance.

To avoid degrading this model's performance, MiniMax highly recommends preserving reasoning between turns.

By:MiniMaxInput Type:Output Type:Publish time:2025-12-22

Providers

Route requests across multiple providers. Copy a provider slug to set your preference.

ProviderCache Hit RateContext
MiniMax
minimax
-$0.3/ M tokens$1.2/ M tokensRead:0.03/ M tokensWrite:0.375/ M tokens204.8K--

Uptime

Direct request success rate on AI Gateway and per-provider.

Throughput

P50 throughput on live AI Gateway traffic, in tokens per second (TPS).

Latency

P50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds.

Activity

Token volume and request traffic to this model over time.

Token Consumption

Related Models

More models from MiniMax