Call MiniMax-M3.1-Flash-Preview through a single OpenAI-compatible endpoint — one key, no separate vendor account, no multi-provider wiring.
MiniMaxChatReasoningCodingVisionPrices are per 1M tokens in USD · ¥6.5 = $1 · Billed in CNY
MiniMax models are built for throughput. The highspeed variants are tuned for very low latency, which makes them a good fit for interactive products where perceived speed matters more than the last point of benchmark score.
At $0.65 per 1M input tokens it sits in line with the catalogue median of $0.65 — a balanced choice. Best for coding agents and repository-scale edits. Every model here is reachable with the same API key, so trialling it against a stronger sibling costs one line of code.
from openai import OpenAI
client = OpenAI(
base_url="https://starseaapi.com/v1",
api_key="sk-your-key"
)
r = client.chat.completions.create(
model="MiniMax-M3.1-Flash-Preview",
messages=[{"role": "user", "content": "Explain prefix caching in one paragraph"}]
)
print(r.choices[0].message.content)| Models | Input $/M | Output $/M | Intelligence index | Coding index |
|---|---|---|---|---|
| embo-01 | $0.08 | $0.08 | — | — |
| MiniMax-M2 | $0.32 | $1.29 | — | — |
| MiniMax-M2.1 | $0.32 | $1.29 | — | — |
| MiniMax-M2.5 | $0.32 | $1.29 | — | — |
| MiniMax-M2.7 | $0.32 | $1.29 | — | 52.6 |
| MiniMax-M2.1-highspeed | $0.65 | $2.58 | — | — |
| MiniMax-M2.5-highspeed | $0.65 | $2.58 | — | — |
| MiniMax-M2.7-highspeed | $0.65 | $2.58 | — | — |
| MiniMax-M3 | $0.65 | $2.58 | 35.7 | 58.6 |