The Rundown AI homepage

Independent tool overview

MiniMax-M2.1 at a glance

MiniMax-M2.1 is an open-weight mixture-of-experts model built for multilingual coding, tool use and long-running agents; it remains available but has been superseded by M2.7 and M3.

Visit the official MiniMax-M2.1 site ↗
MiniMax-M2.1 product preview
Developer
MiniMax
Released
December 23, 2025
Architecture
230B MoE, about 10B active
API context
204,800 tokens total
Current status
Supported legacy model
Current successors
MiniMax-M2.7 and M3

Overview

What MiniMax-M2.1 is

MiniMax released M2.1 in December 2025 as a coding- and agent-focused update to M2. The 230-billion-parameter mixture-of-experts model activates roughly 10 billion parameters per inference and emphasizes non-Python development, including Rust, Java, Go, C++, Kotlin, Objective-C, TypeScript and JavaScript. It also targeted native iOS and Android work, visual web development, tool calling and multi-step office tasks.

M2.1 is no longer MiniMax's current flagship. The API still supports both standard and high-speed endpoints, but pricing documentation places them under legacy models. MiniMax-M2.7 is the current M2-series text endpoint, while M3 adds a one-million-token context window, native image and video input and computer use. Use M2.1 when compatibility, its open weights or an existing evaluation requires it; start a new project by testing the successors first.

Use cases

Who MiniMax-M2.1 is best for

The strongest fit depends on the job you need the product to complete, not the size of its feature list.

Existing M2.1 integrations

Keep a pinned model stable while testing behavior, latency and cost against newer MiniMax endpoints.

Multilingual software repositories

Apply the model to codebases that mix systems, backend, web and native mobile languages rather than only Python.

Agent-scaffold research

Evaluate tool use and context-management behavior across Claude Code-compatible, Cline and other agent environments.

Organizations needing local weights

Run the published model under its modified MIT license when the infrastructure and attribution requirement are acceptable.

Capabilities

Core MiniMax-M2.1 features

1

Polyglot coding

Training and evaluation focused on more than ten programming languages across systems, backend, web and native application work.

2

Web and mobile app development

M2.1 was tuned for frontend aesthetics, complex interactions and both Android and iOS implementation tasks.

3

Agent and tool use

Use function calling, long-horizon planning and common agent instruction formats such as CLAUDE.md, AGENTS.md-style files, skills and slash commands.

4

Interleaved thinking

The model can reason between tool calls during multi-step tasks, with MiniMax recommending compatible context handling for best results.

5

Standard and high-speed APIs

Choose the regular endpoint at roughly 60 output tokens per second or the higher-priced high-speed endpoint at roughly 100.

6

Open weights

Download the 230 GB model release and serve it through frameworks including SGLang, vLLM, Transformers, MLX-LM or KTransformers.

Process

How the MiniMax-M2.1 workflow works

  1. Step 1

    Decide whether M2.1 is actually required

    For a new integration, benchmark M2.7 and M3 first; keep M2.1 for compatibility, reproducibility or a validated cost-quality advantage.

  2. Step 2

    Select API or self-hosting

    Use MiniMax's API for lower operational burden, or deploy weights when infrastructure control justifies the model's size and serving complexity.

  3. Step 3

    Configure the agent correctly

    Preserve tool results and relevant reasoning context, provide concise repository instructions and define explicit success checks.

  4. Step 4

    Run real repository evaluations

    Measure correctness, tests passed, latency, token cost and human review effort on representative tasks rather than relying only on vendor benchmarks.

  5. Step 5

    Pin and monitor the model

    Use an explicit model ID, watch legacy-support notices and maintain a tested migration path to a current endpoint.

Cost

MiniMax-M2.1 pricing and free plan

MiniMax still offers M2.1 through pay-as-you-go API endpoints, but lists it under legacy models. Standard costs $0.30 per million input tokens and $1.20 per million output tokens; Highspeed doubles those rates. Prompt-cache reads are $0.03 per million tokens and writes are $0.375 per million on both endpoints.

MiniMax-M2.1 API

$0.30 input / $1.20 output per 1M tokens

The standard legacy endpoint at approximately 60 output tokens per second.

  • $0.03 per million cached-read tokens
  • $0.375 per million cache-write tokens
  • 204,800-token total context
  • Pay-as-you-go billing

MiniMax-M2.1-highspeed API

$0.60 input / $2.40 output per 1M tokens

The same model behavior served at approximately 100 output tokens per second.

  • $0.03 per million cached-read tokens
  • $0.375 per million cache-write tokens
  • 204,800-token total context
  • Pay-as-you-go billing

Open weights

No model download fee

Self-host under MiniMax's modified MIT license and pay the resulting infrastructure and operations cost.

  • Approximately 230 GB model repository
  • Commercial products must prominently display “MiniMax M2.1” in the user interface
  • Serving hardware, storage and engineering not included
  • License and third-party dependencies require review

Pricing checked . Check current pricing at the source ↗

Assessment

MiniMax-M2.1 strengths and limitations

Where it stands out

  • Strong emphasis on multilingual, real-world software development
  • Standard API pricing remains low relative to many hosted frontier models
  • High-speed endpoint provides a clear latency option without changing model behavior
  • Open weights enable private deployment and reproducible evaluation
  • Broad agent-scaffold and tool-calling support
  • The API remains available even after newer models launched

What to consider

  • MiniMax now categorizes M2.1 as a legacy model, so new systems should compare M2.7 and M3 before adopting it
  • The 204,800-token context is far below M3's advertised one-million-token window
  • M2.1 is text-focused and lacks M3's native image, video and computer-use capabilities
  • The roughly 230 GB weight release makes practical self-hosting expensive and operationally complex
  • The modified MIT license requires prominent “MiniMax M2.1” attribution in the user interface of commercial products
  • Vendor benchmark scores depend on specific agent scaffolds and should be validated on the buyer's repositories
  • Generated code and tool actions still require tests, security review and human approval

Compare

MiniMax-M2.1 alternatives

The right alternative depends on the specific output, workflow, controls and budget your project requires.

Consumer

Minimax M2.7

Choose MiniMax-M2.7 for the current M2-series coding and agent endpoint at the same standard input and output rates.

Explore Minimax M2.7

Consumer

MiniMax M3

Choose MiniMax M3 when one-million-token context, multimodal input or computer use matters.

Explore MiniMax M3

Coding

Claude Code

Consider Claude Code when the priority is a mature coding-agent product rather than self-hosting an open-weight model.

Explore Claude Code

Questions

MiniMax-M2.1 FAQs

Is MiniMax-M2.1 still available?

Yes. MiniMax still lists standard and high-speed M2.1 API endpoints, but its pricing page categorizes them as legacy models. The open weights also remain available.

What replaced MiniMax-M2.1?

MiniMax released M2.5 and then M2.7 in the text-model line. M3 is the newer flagship family with one-million-token context, native multimodality and computer use.

How much does the MiniMax-M2.1 API cost?

Standard M2.1 costs $0.30 per million input tokens and $1.20 per million output tokens. The high-speed endpoint costs $0.60 input and $2.40 output per million tokens.

Can MiniMax-M2.1 be self-hosted?

Yes. MiniMax publishes the weights and deployment guides for SGLang, vLLM, Transformers and other frameworks. The model repository is roughly 230 GB, so production serving still requires substantial infrastructure.

Can MiniMax-M2.1 be used commercially?

The modified MIT license permits commercial use, but a commercial product or service must prominently display “MiniMax M2.1” in its user interface. Review the full license before deployment.

Bottom line

Our MiniMax-M2.1 verdict

MiniMax-M2.1 remains a capable, low-cost and deployable coding model, but it is now a compatibility choice rather than the default starting point. Existing users can keep it pinned while evaluating migrations; new users should benchmark M2.7 and M3 first, then choose M2.1 only if its open weights, behavior or latency profile produces a measurable advantage.

Visit MiniMax-M2.1 website ↗
The Rundown University

AI training for the future of work.

Get access to all our AI courses, hundreds of real-world AI use cases, live expert-led workshops, an exclusive network of AI early adopters, and more.

AI Courses

Get unlimited access to all of our current & upcoming industry-specific AI courses for the duration of your subscription.

Daily Guides

To keep up with the rapid pace of AI, our team publishes AI implementation guides daily. Our library contains 300+ practical use cases to automate real-world work.

Workshops

Join weekly, live, interactive sessions with industry leaders who are at the forefront of AI for hands-on implementation guidance and exclusive insights.

Community

Network with an exclusive community of AI-first professionals who are working smarter with AI. Learn how early adopters are using AI in their work and businesses.