MiniMax M2.1 Highspeed API

minimax/m2-1-highspeed
Unlike traditional large language models optimized primarily for depth, MiniMax M2.1 Highspeed prioritizes computational efficiency without sacrificing coherence, contextual understanding, or instruction adherence.
Context
200K tokens
Input
$0.78 / 1M tokens
Output
$3.12 / 1M tokens
Released
Apr 14, 2026

How to use MiniMax M2.1 Highspeed API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to minimax/m2-1-highspeed.
import requests

r = requests.post(
    "https://api.aimlapi.com/v1/chat/completions",
    headers={"Authorization": "Bearer " + AIMLAPI_KEY},
    json={
      "model": "minimax/m2-1-highspeed",
      "messages": [
        {
          "role": "user",
          "content": "Hello!"
        }
      ]
    },
)
print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "minimax/m2-1-highspeed",
    "messages": [
      {
        "role": "user",
        "content": "Hello!"
      }
    ]
  }),
});
console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"minimax/m2-1-highspeed","messages":[{"role":"user","content":"Hello!"}]}'

OpenAI-compatible — swap the base URL and it works with your existing SDK.

MiniMax M2.1 Highspeed API Pricing

TypePrice
Input
$0.78 / 1M tokens
Output
$3.12 / 1M tokens
Cached input
$0.78 / 1M tokens

MiniMax M2.1 Highspeed vs other models

ModelInputOutputContextBest for
$0.78 / 1M tokens
$3.12 / 1M tokens
200K tokens
Coding + agents
$6.5 / 1M tokens
$39 / 1M tokens
1M tokens
Reasoning + agents
$2.6 / 1M tokens
$13 / 1M tokens
1M tokens
Balanced coding + agents
$3.9 / 1M tokens
$19.5 / 1M tokens
1M tokens
Long-context, multimodal & agentic workflows
$0.65 / 1M tokens
$3.9 / 1M tokens
1M tokens
Reasoning + agents

Frequently asked questions

MiniMax M2.1 Highspeed has a 204,800 tokens context window and can return up to 204,800 tokens.

MiniMax M2.1 Highspeed takes image, text as input and returns text.

Use minimax/m2-1-highspeed as the model id. Requests go to https://api.aimlapi.com/v1/chat/completions.

MiniMax M2.1 Highspeed is priced at input $0.78 / 1M tokens, output $3.12 / 1M tokens, cached input $0.78 / 1M tokens.

Yes, MiniMax M2.1 Highspeed can stream responses as they are generated.

Yes, MiniMax M2.1 Highspeed accepts image input alongside text.

MiniMax M2.1 Highspeed was built by MiniMax.

Yes, it supports tool use, parallel tool calls, and function calling.

Start building with MiniMax M2.1 Highspeed

Get API Key
1000+ models, one API.