MiniMax M3

MiniMax-M3

MiniMax M3 is a frontier multimodal MiniMax model with a 1M-token context window, built for long-horizon agent workflows, coding, and tool use. It uses MiniMax Sparse Attention to reduce the cost of large-context workloads compared with earlier generations. Use the MiniMax API through TokenHub to review the current MiniMax M3 price, model ID, supported endpoints, context limits, and benchmark results before deploying it for production software tasks or collaborative agent workflows.

Context Window

512K tokens

Maximum Output

128K tokens

Release Date

Jun 1, 2026

Modalities

MiniMax M3 Pricing

Input PriceOutput PriceCache Read
$0.6/M$2.4/M$0.12/M

MiniMax M3 API Capabilities

Reasoning

Supported

Tool calling

Supported

Temperature parameter

Supported

Attachments

Supported

Knowledge Base

—

Endpoint Protocols

Completions APIMessages APIgemini

MiniMax M3 Model Highlights

MiniMax M3 combines advanced software engineering, a million-token context window and native visual understanding for long-running development and agent tasks.

Long-Horizon Coding

Handles bug fixing, frontend and backend development, performance optimization and repeated collaboration across extended engineering sessions.

Million-Token Context

MiniMax Sparse Attention enables context windows up to one million tokens while reducing the computational burden of long-context processing.

Native Multimodality

Processes text, images and video through a model trained with mixed modalities from the beginning, supporting visually grounded reasoning.

MiniMax M3 Use Cases

MiniMax M3 suits repository-scale engineering, long-running technical research and visual analysis that combines documents, code and experimental results.

Repository Engineering

Analyze a large codebase, discuss evolving requirements and implement coordinated frontend, backend or performance changes over multiple rounds.

Technical Research Reproduction

Keep a paper, source code and experiment logs in one context, then implement experiments, inspect figures and compare results with the publication.

Multimodal Technical Analysis

Interpret diagrams, charts, screenshots or video together with technical text and code to identify issues and propose implementation changes.

How to Use MiniMax M3 via the TokenHub API

Create API key
import OpenAI from "openai"

const client = new OpenAI({
  apiKey: process.env.TOKENHUB_API_KEY,
  baseURL: "https://us-api.tokenhub.com/v1",
})

const result = await client.chat.completions.create({})
console.log(result.choices[0]?.message?.content)

MiniMax M3 Benchmarks

MiniMax-M3

Index score
Artificial Analysis Intelligence IndexArtificial Analysis broad capability aggregate44.4
Artificial Analysis Coding IndexArtificial Analysis software task aggregate43.4
Knowledge & Reasoning
GPQAAdvanced science problem solving92.9%
HLEBroad expert-level exam set37.1%
Coding & Engineering
SciCodeScientific coding challenges45.4%
Terminal-Bench HardHard terminal task execution42.4%
Instruction Following & Agent Tasks
IFBenchPrompt constraint adherence82.9%
AA-LCRLong-context reasoning74%
τ²-BenchAgent workflow tasks88.9%

Metrics sourced from Artificial Analysis

Media and Discussions

Selected public videos and posts related to this model.

X (Twitter)

View post on X
View post on X
View post on X

Reddit

YouTube

Watch on YouTube
Watch on YouTube
Watch on YouTube

MiniMax M3 API FAQs

Answers to common questions about the MiniMax M3 API, pricing, coding and agent workflows, long-context limits, and model comparisons on TokenHub.

What is MiniMax M3?+

MiniMax M3 is a MiniMax multimodal model for long-horizon agent work, coding, and tool use. Its model profile emphasizes large-context workloads and production-oriented software tasks where a model needs to work across substantial project information.

Is MiniMax M3 suitable for coding and agent workflows?+

MiniMax M3 is worth evaluating for coding and agent workflows that need long context, repeated tool use, or collaboration across multiple task steps. Before adopting it in production, test it with representative repository tasks, tool calls, and completion criteria from your own workflow.

How do I use the MiniMax M3 API through TokenHub?+

Create a TokenHub API key, copy the exact MiniMax M3 model ID shown on this page, and choose a supported endpoint in the API section. Use the generated request example as the starting point for your OpenAI-compatible client or other supported integration.

How should I evaluate MiniMax M3 price?+

Review the current input, output, and cache-read prices on this page, then estimate cost using your typical prompt size, response length, request volume, and cache-hit rate. Long-context and agent workflows can have a very different cost profile from a short text-only request.

What should I check before using MiniMax M3 for a long-context task?+

Check the current context window, maximum output, input modalities, supported endpoints, and pricing shown in the model specifications. A large context limit does not remove the need to structure documents, control tool loops, and test how the model performs with your real project data.

MiniMax M3 vs Kimi K3: which model should I choose?+

Compare MiniMax M3 and Kimi K3 against the workload you need to ship. Review the current API price, context and output limits, modalities, supported endpoints, and benchmark data, then run a small evaluation using your own coding, reasoning, or agent tasks before committing to either model.

Where can I find a MiniMax API key and the current MiniMax M3 model ID?+

Create an API key in the TokenHub workspace, then copy the exact MiniMax M3 model ID displayed on this page. Confirm the supported endpoint and current specifications here before sending production traffic, because availability and model details can change.

Ready to use MiniMax M3?

Use one API key to access MiniMax M3 and more AI models through TokenHub.

Create API key