MiniMax M3 API

Call MiniMax M3 through Wokey and review published catalog pricing, upstream-native API forms, context, and model capabilities.

Model ID
MiniMax-M3
Vendors
minimax, volcengine, opencode
Supply status
Currently callable

Catalog pricing snapshot

Prices are per 1M tokens; request-time pricing controls settlement.

MeterWokeyOfficial referenceSavings
Input$0.09$0.4580%
Output$0.36$1.880%
Cache read$0.018$0.0980%
Cache write———

Official price source: https://platform.minimax.io/docs/guides/pricing-paygo

Model capabilities

Context
1,000,000 tokens
Max output
80,000 tokens
Streaming
Supported
Tools
Supported
Vision
—
Supported API forms
/v1/chat/completions

Send a request

/v1/chat/completions

curl https://api.wokey.ai/v1/chat/completions \
  -H "Authorization: Bearer $WOKEY_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"MiniMax-M3","messages":[{"role":"user","content":"Hello"}]}'

Use MiniMax M3 in your tools

The gateway is https://api.wokey.ai (OpenAI-compatible clients usually take https://api.wokey.ai/v1), and the model ID is MiniMax-M3. Open the guide for the client you use.

Frequently asked questions

How much does the MiniMax M3 API cost, and how much does it save versus the official API?

Wokey lists input at $0.09 per 1M tokens and output at $0.36 per 1M tokens; the official references are $0.45 and $1.8. That is 80% lower for input and 80% lower for output. Request-time pricing controls settlement, and the vendor pricing source is linked on this page.

What is the MiniMax M3 API model ID?

The model ID for MiniMax M3 on Wokey is MiniMax-M3. Use it exactly as written in the request body's model field or in your client's model setting (for example ANTHROPIC_MODEL in Claude Code, or model in the Codex CLI config.toml); GET https://api.wokey.ai/v1/models also lists it.

Which API forms does MiniMax M3 support, and how do I call it through Wokey?

MiniMax M3 supports /v1/chat/completions (Chat Completions requests). Set the API base URL to https://api.wokey.ai, authenticate with a Wokey API key, and use the canonical model ID MiniMax-M3; you can start with /v1/chat/completions. The Send a request section includes a curl example for every supported API form.

How do I use MiniMax M3 in opencode, Hermes Agent, and similar clients?

These clients call MiniMax M3 over OpenAI Chat Completions: point the API address at https://api.wokey.ai (clients that expect an OpenAI SDK-style address take https://api.wokey.ai/v1), add your Wokey API key, and use the model ID MiniMax-M3. Each client's guide shows its config format.

What are the context window and maximum output for MiniMax M3?

The catalog context window is 1,000,000 tokens and the maximum output is 80,000 tokens. These are separate limits, and a request must also satisfy the upstream constraints in effect when it is processed.

How are cache reads and cache writes priced for MiniMax M3?

Cache read: Wokey lists $0.018 per 1M tokens and the official reference is $0.09 per 1M tokens. Cache write: the vendor does not publish a separate meter, so Wokey also shows — as the public price.

How does calling MiniMax M3 through Wokey differ from OpenRouter?

Wokey lists MiniMax M3 at $0.09 input / $0.36 output per 1M tokens, against an official reference of $0.45 / $1.8. OpenRouter states it adds no markup on inference and charges its fee when you buy credits (5.5% by card). Wokey carries a curated set of models and has no per-provider routing parameters, so OpenRouter fits better if you need a wider catalog. The OpenRouter alternative page has the full comparison.

Does MiniMax M3 support streaming, tool calling, and image input?

Streaming is supported; tool calling is supported; image input is not declared in the catalog. These values come directly from the current runtime model catalog and are not inferred from the model name.

Related models and documentation

Get API key · MiniMax-M3