MiMo V2.5

mimo-v2.5

Xiaomi MiMo V2.5 is Xiaomi’s native multimodal Mixture-of-Experts model for production AI applications. It understands text, images, video, and audio, and generates text for chat, coding, reasoning, content generation, and agent workflows. With long-context and tool-use capabilities, MiMo V2.5 is suited to document analysis, multimodal assistants, structured extraction, and complex tasks that combine multiple input formats. Use the Xiaomi MiMo API through TokenHub to review the current MiMo V2.5 price, model ID, supported endpoints, and model limits before production use.

Context Window

1.1M tokens

Maximum Output

131.1K tokens

Release Date

Apr 23, 2026

Modalities

MiMo V2.5 Pricing

Input PriceOutput PriceCache Read
$0.14/M$0.28/M$0.0028/M

MiMo V2.5 API Capabilities

Reasoning

Supported

Tool calling

Supported

Temperature parameter

Supported

Attachments

Supported

Knowledge Base

—

Endpoint Protocols

Completions APIResponses APIMessages API

MiMo V2.5 Model Highlights

MiMo V2.5 combines native full-modal understanding, general-purpose agent capabilities, and a one-million-token context for responsive multimodal applications.

Native Full-Modal Understanding

The model jointly understands text, images, video, and audio, enabling information from different media to be considered within one task.

General Agent Capability

MiMo V2.5 can translate multimodal perception into actions for everyday agent tasks involving planning and tool coordination.

Responsive Million-Token Context

A one-million-token context and higher average inference speed support large inputs while preserving responsiveness for frequent tasks.

MiMo V2.5 Use Cases

MiMo V2.5 is suited to multimedia analysis, chart and video understanding, and everyday agents that combine perception with practical actions.

Multimedia Content Analysis

Combine text, images, audio, and video to identify events, topics, speakers, and relationships, then return a unified text summary.

Chart and Video Understanding

Interpret dense charts or lengthy video material, extract important changes and evidence, and answer questions about the content.

Multimodal Everyday Agents

Understand mixed-media requests, plan a short sequence of actions, coordinate tools, and return a practical result for routine work.

How to Use MiMo V2.5 via the TokenHub API

Create API key
import OpenAI from "openai"

const client = new OpenAI({
  apiKey: process.env.TOKENHUB_API_KEY,
  baseURL: "https://us-api.tokenhub.com/v1",
})

const result = await client.chat.completions.create({})
console.log(result.choices[0]?.message?.content)

MiMo V2.5 Benchmarks

MiMo-V2.5

Index score
Artificial Analysis Intelligence IndexArtificial Analysis broad capability aggregate38
Artificial Analysis Coding IndexArtificial Analysis software task aggregate56.8

Metrics sourced from Artificial Analysis

MiMo V2.5 Comparisons

Xiaomi MiMo V2.5 Media and Discussions

Selected public discussions and videos about Xiaomi MiMo V2.5 and the MiMo V2.5 model family.

X (Twitter)

View post on X

Reddit

YouTube

Watch on YouTube

Xiaomi MiMo V2.5 API FAQs

Answers to common questions about the Xiaomi MiMo V2.5 API, pricing, multimodal capabilities, coding and agent workflows, and model comparisons on TokenHub.

What is Xiaomi MiMo V2.5?+

Xiaomi MiMo V2.5 is Xiaomi’s native multimodal Mixture-of-Experts model for production AI applications. It is intended for tasks that combine text generation with multimodal understanding, coding, reasoning, content generation, and agent workflows.

What can the Xiaomi MiMo V2.5 API be used for?+

The Xiaomi MiMo V2.5 API can be evaluated for AI assistants, content generation, document analysis, structured extraction, coding support, and agent workflows that use multiple input formats. Confirm the supported modalities, endpoint, and request schema on the current model page before implementation.

Does Xiaomi MiMo V2.5 support multimodal input and tool calling?+

MiMo V2.5 is positioned as a native multimodal model with tool-use capabilities. Check the modalities, tool-calling support, and available endpoints shown on this TokenHub model page before relying on a specific input type or function in production.

How do I use the Xiaomi MiMo V2.5 API through TokenHub?+

Create a TokenHub API key, copy the exact Xiaomi MiMo V2.5 model ID shown on this page, and choose a supported endpoint in the API section. Use the generated request example as the starting point for your OpenAI-compatible client or other supported integration.

How should I evaluate Xiaomi MiMo V2.5 price?+

Review the current input, output, and cache-read prices displayed on this page, then estimate cost using your typical prompt size, response length, request volume, and cache-hit rate. For multimodal and agent workloads, also test how input format and tool loops affect total usage.

MiMo V2.5 vs Xiaomi MiMo V2.5 Pro: which should I choose?+

Choose based on the workload, not the suffix. Compare current API pricing, context and output limits, modalities, tool-calling support, latency, and results from your own evaluation set. Use the MiMo V2.5 vs MiMo V2.5 Pro comparison page to verify live specifications before standardizing on either model.

Where can I find a Xiaomi MiMo API key and the MiMo V2.5 model ID?+

Create an API key in the TokenHub workspace and copy the exact Xiaomi MiMo V2.5 model ID displayed on this page. Confirm the supported endpoint and current specifications before sending production traffic, because availability and model details can change.

Ready to use MiMo V2.5?

Use one API key to access MiMo V2.5 and more AI models through TokenHub.

Create API key