← All open-weight models

Chinese AI labs · General language

MiniMax-M2.7-highspeed

A 2x latency surcharge on identical weights — justify it with a measured p95 requirement, not a hunch.

Open weights · Community or custom license · Data checked September 6, 2026 · Not a hands-on evaluation

Model ID MiniMax-M2.7-highspeed

Parameters
229B
Architecture
Not recorded
Context window
204,800 tokensMaximum input
Max output
Not recordedPer response
Input price
$0.60 per 1M tokens
Output price
$2.40 per 1M tokens
Cached input
$0.06 per 1M tokens
Blended price
$1.05 per 1M tokens3:1 input to output
Knowledge cutoff
Not recorded
Released
Not recorded
Status
ga
Size band
120B and above

License and openness

MiniMax-M2.7-highspeed is released under the MiniMax Model License (custom; see LICENSE on Hugging Face). This is a community or custom license with its own conditions, which can include acceptable-use rules, user thresholds or naming requirements. Read it before commercial use. License terms can change between versions; confirm the text published with the weights you download.

Capabilities

Input modalitiestext
Output modalitiestext

What it is for

At a three-to-one input-to-output ratio, MiniMax-M2.7-highspeed costs $1.05 per million tokens blended. A workload of one million input and 330,000 output tokens per day would run about $41.76 per month at list price, before caching or batch discounts.

Cost at list price

Input$0.60 per 1M tokens
Output$2.40 per 1M tokens
Cached input$0.06 per 1M tokens
Blended, 3:1$1.05 per 1M tokens
Estimate a workload across all models ↗

Sources

Sources checked September 6, 2026. Specifications and prices change without notice; confirm against the provider before you commit.

Other open-weight models from Chinese AI labs

See all 17 from Chinese AI labs ↗

Browse all open-weight models ↗