MiniMax (Reasoning & Code)
MiniMax is a solid Chinese provider, but DeepSeek and MiMo remain cheaper and more capable for most workloads. Consider MiniMax if you want provider diversity or have specific needs its model lineup serves well.
Getting an API Key
- Go to platform.minimaxi.chat
- Sign up or log in
- Navigate to API Keys and create a new key
- Paste it into Wolffish → Settings → Models → MiniMax
Models
Video Models Are Not Chat Models
MiniMax’s H3 video models never appear in the model picker — you don’t chat with them. They are reached through the video capability, which your current chat model calls as a tool.The video key is a separate field. MiniMax issues one credential that unlocks both its chat models and its video API, but Wolffish never copies the value between Settings → Models → MiniMax and Settings → Services → Video generation. Video generation is a service, not a property of whichever brain you’re chatting with — so switching your chat brain away from MiniMax never takes video with it, and rotating one key never breaks the other. Paste the key in both places if you want both.
Reasoning modes
How this model reasons is set in the model card beside the chat input — hover the model switch to preview the card, click to pin it open. The top chip row, labelled Thinking, is the control: pick a chip and it applies from your next message. Two separate ideas combine here:Thinking — whether the model reasons
- Off — the model answers immediately. Fastest and cheapest; ideal for simple, direct tasks.
- On (the chip reads Normal) — the model first works through the problem in a dedicated reasoning pass before replying. Slower and uses more tokens, but markedly more accurate on multi-step, logical, or ambiguous tasks.
Effort — how hard it thinks
Only effort-capable models expose this; it applies once thinking is on.- High — standard reasoning depth. The right default for most agentic work.
- Max — the model reasons longer and deeper for the hardest problems. More tokens and latency in exchange for higher quality on complex work.
The chip row
The active chip is tinted; the rest sit quiet. Each model shows only the chips it genuinely supports — a model that always reasons has no Off, a model with a single mode shows that one chip inert, and a model that can’t reason at all replaces the row with “Reasoning is not supported by this model.” Wolffish remembers your choice per model. (Before v1.0.236 this was a colour-coded brain button on the composer; it moved into the model card, beside the Single/Workflow row, so every model knob sits in one panel.)
On MiniMax: MiniMax-M3 is On / Off. The M2.x models reason always-on, so the button is locked on.