MiniMax M3.1 Flash API pricing and integration guide
Review MiniMax M3.1 Flash's published 1x rate, 1M context, and image support.
Current prices and model reference
- Model ID
MiniMax-M3.1-flash- Documented context
- 1M
- Documented image input
- Supported
- Live model multiplier
- 1×
- Input price
- ¥0.4 / 1M tokens
- Output price
- ¥2 / 1M tokens
Prices are in CNY. A 1× multiplier equals CNY 0.4 per million input tokens. Output and cache charges follow the billing configuration; your usage record is authoritative.
View the full live pricing catalogWhen to consider this model
Consider it for everyday coding, batch tasks, and cost-sensitive agents. Compare stability and retry cost on a small sample first.
How to evaluate it
Test text and screenshots with known answers, then increase concurrency gradually while tracking latency, failures, and per-task cost.
What to check before integration
Use the exact MiniMax-M3.1-flash ID; the new route is listed in the live catalog, and real requests confirm final availability.
Client configuration
In a client supporting a custom API endpoint, enter this base URL and model ID with your own API key. Select the protocol described in the client tutorial and validate a short request first.
Base URL: https://api.llm-token.cn/v1
Model ID: MiniMax-M3.1-flash