MiniMax M3 is a frontier model optimized for agentic coding workflows and long-context multimodal analysis. Developers building autonomous agents or debugging complex repositories should use it when they need reliable tool execution across a million-token window.
Modalities
Text+Image → Text
Input
$0.3852 /1M
Output
$1.54 /1M
Cached input
$0.03852 /1M90% off
Context
1M
Speed
medium
Performance
Live production data from real requests on Kyma — not synthetic benchmarks.
Rank
#34
of 87 active models
Tokens served
379.5K
all-time
Success rate
100%
last 7 days
Pricing
Pay per token. Cached input is billed at 10% of the input rate.
+ Estimate your workload− Estimate your workload
Estimated monthly cost
$39.97
$1.33 / day on MiniMax M3
Same workload on:
Estimates use list pricing with cached input billed at the 90%-discount rate. Actual bills depend on real token counts — every response includes its exact cost.
When to use MiniMax M3
Where this model earns its cost — and where it doesn't.
MiniMax M3 uses MSA sparse attention to process text, image, and video inputs within a 1,048,576-token context window. It supports tool calling, structured outputs, and explicit reasoning traces, making it suitable for multi-step coding tasks and repository-level analysis.
On Kyma, the model runs through an OpenAI-compatible endpoint with automatic request failover and exact cost reporting in the usage.cost field. Prompt caching applies a 90% discount to repeated prefixes, and new accounts receive a $0.50 credit to test throughput.
Output generation caps at 32,768 tokens, and median throughput sits around 2 tokens per second. It is not designed for high-speed conversational streaming or low-latency real-time applications.
Agentic Code Generation
Handles multi-step repository edits and tool chaining across long context windows.
Multimodal Input Analysis
Processes text, image, and video inputs to extract structured data or debug visual workflows.
Long-Horizon Debugging
Maintains state over extended sessions to trace errors across large codebases.
Structured Output Parsing
Returns deterministic JSON or XML formats for downstream pipeline integration.
Not ideal for: Avoid this model for low-latency chat or high-throughput streaming where sub-second response times are required.
How it compares
Against the peers people actually weigh it against.
| Spec | MiniMax M3 | MiniMax M2.5 | MiniMax M2.7 |
|---|---|---|---|
| Input /1M | $0.3852 | $0.3826 | $0.405 |
| Output /1M | $1.54 | $1.346 | $1.62 |
| Context | 1M | 197K | 205K |
| Tools | Yes | Yes | Yes |
| Reasoning | Yes | Yes | Yes |
| Speed | medium | medium | medium |
Quick start
Up and running in under two minutes.
- 1
Create an API key
Sign up and grab a key from the dashboard — $0.50 free credit, no card required.
Get API key → - 2
Make your first request
Drop in your key and send a chat completion — fully OpenAI-compatible.
curl https://kymaapi.com/v1/chat/completions \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "minimax-m3", "messages": [ {"role": "user", "content": "Explain prompt caching in one paragraph."} ] }' - 3
Stream responses
Add
"stream": trueto receive tokens as they arrive.curl https://kymaapi.com/v1/chat/completions \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "minimax-m3", "stream": true, "messages": [{"role": "user", "content": "Hello!"}] }'
FAQ
Common questions about this model.
What is the context window of MiniMax M3?
How much does the MiniMax M3 API cost?
Does MiniMax M3 support function calling?
How do I use MiniMax M3?
Does it support image and video inputs?
How does prompt caching work on Kyma?
Can I rely on it for structured JSON outputs?
More models by MiniMax
See all 13 →| Model | Context | Input | Output |
|---|---|---|---|
MiniMax Music Pro | — | $0.21 / song | |
MiniMax M2.7 | 205K | $0.405 | $1.62 |
MiniMax M2.5 | 197K | $0.3826 | $1.346 |
MiniMax Music | — | $0.045 / song | |
MiniMax Voice Design | — | $4.20 / call | |
Hailuo 02 (1080p) | — | $0.78 / video | |
Hailuo 02 (768p) | — | $0.42 / video | |
Hailuo 02 (512p) | — | $0.14 / video | |
