# MiniMax M2.7 HighSpeed

> MiniMax M2.7 HighSpeed is the latency-oriented variant of M2.7. It carries the same 204K-token context window, the same 131K output headroom, the same text-only input and the same capability set — reasoning, coding, tool calling, prompt caching — at twice the token price.

Canonical: https://routemux.com/minimax/minimax-m2.7-highspeed
Provider: MiniMax

## Key facts

| Field | Value |
| --- | --- |
| Model ID | minimax/minimax-m2.7-highspeed |
| Official model ID | MiniMax-M2.7-highspeed |
| Context window | 205K |
| Max output | 131K |
| Input modalities | text |
| Output modalities | text |
| Capabilities | Reasoning, Tool Calling, Coding, Cache |
| Availability | available |

## Pricing

- RouteMux input: $0.06 per 1M tokens
- RouteMux output: $0.24 per 1M tokens
- Official input: $0.6 per 1M tokens
- Official output: $2.4 per 1M tokens
- Cache read: $0.006 per 1M tokens
- Saving vs official: 90%
- Official price source: minimax_text_x0.10, verified 2026-08-07

Billing is prepaid wallet, charged per successful request only.

## API access

Base URL: https://api.routemux.com

Supported protocols:

- openai_chat
- openai_responses
- anthropic_messages

## What is MiniMax M2.7 HighSpeed

MiniMax M2.7 HighSpeed is the latency-oriented variant of M2.7. It carries the same 204K-token context window, the same 131K output headroom, the same text-only input and the same capability set — reasoning, coding, tool calling, prompt caching — at twice the token price.

Through RouteMux it is $0.06 per 1M input tokens against MiniMax's $0.60, and $0.24 against $2.40 on output.

Since the structural specs are identical to M2.7, the only reason to pay double is that response latency is the constraint. If it is not, M2.7 does the same work for half.

## Key capabilities

- **Same specs as M2.7** — 204K context, 131K output, text-only, same capability set. The variable is latency, not capability.
- **Twice the token price** — $0.06 vs $0.03 input — pay it only when latency is the binding constraint.
- **Tool calling and caching** — Reasoning, coding, tool calling and prompt caching are supported.
- **Three protocols** — OpenAI Chat, OpenAI Responses and Anthropic Messages.

## Other variants in this family

- MiniMax M2.7 (`minimax/minimax-m2.7`)

## Sources

- [MiniMax — API docs](https://platform.minimaxi.com/document)

_Content reviewed by Jerry Fan on 2026-08-09._
