o3-mini

o3-mini

o3 Mini is the cost-efficient member of OpenAI’s o-series reasoning models. The launch materials focus on STEM tasks such as science, math, and coding, with lower latency and lower cost than larger reasoning models. It is a strong option when users need reasoning behavior but cannot spend the time or budget required by a full o-series model.

Context Window

200K tokens

Maximum Output

100K tokens

Release Date

Dec 20, 2024

Modalities

o3-mini Pricing

Input PriceOutput PriceCache Read
$1.1/M$4.4/M$0.55/M

o3-mini API Capabilities

Reasoning

Supported

Tool calling

Supported

Temperature parameter

Not supported

Attachments

Not supported

Knowledge Base

2024-05-01

Endpoint Protocols

Completions APIMessages APIgemini

o3-mini Model Highlights

o3-mini provides compact text reasoning, long-context processing and structured task execution for focused analytical workloads.

Compact Reasoning

Offers reasoning intelligence in a smaller model for tasks that need analytical depth with lower latency and cost targets.

Long Text Context

Supports a 200,000-token context window for reasoning over lengthy text, code and structured records.

Structured Output Control

Can organize analytical results into consistent schemas, supporting reliable downstream processing of text-based tasks.

o3-mini Use Cases

o3-mini fits text-based reasoning workloads such as policy analysis, text-to-SQL conversion and structured entity extraction.

Policy Analysis

Reads lengthy policies, evaluates a case against stated conditions and explains the resulting decision.

Text-to-SQL Conversion

Translates analytical questions and database schemas into SQL queries, then explains the query logic for review.

Entity Relationship Extraction

Identifies entities and their connections in large text collections and returns records suitable for graphs or databases.

How to Use o3-mini via the TokenHub API

Create API key
import OpenAI from "openai"

const client = new OpenAI({
  apiKey: process.env.TOKENHUB_API_KEY,
  baseURL: "https://us-api.tokenhub.com/v1",
})

const result = await client.chat.completions.create({})
console.log(result.choices[0]?.message?.content)

o3-mini Benchmarks

Index score
Artificial Analysis Intelligence IndexArtificial Analysis broad capability aggregate18.4
Artificial Analysis Coding IndexArtificial Analysis software task aggregate17.3
Knowledge & Reasoning
MMLU-ProAdvanced multi-task knowledge80.2%
GPQAAdvanced science problem solving77.3%
HLEBroad expert-level exam set12.3%
Coding & Engineering
LiveCodeBenchLive coding problems73.4%
SciCodeScientific coding challenges39.8%
Terminal-Bench HardHard terminal task execution6.1%
Math
MATH-500Advanced math problem solving98.5%
AIMECompetition math problems86%
Instruction Following & Agent Tasks
IFBenchPrompt constraint adherence67.1%
AA-LCRLong-context reasoning39.3%
τ²-BenchAgent workflow tasks31.3%

Metrics sourced from Artificial Analysis

Media and Discussions

Selected public videos and posts related to this model.

X (Twitter)

View post on X
View post on X
View post on X

Reddit

YouTube

Watch on YouTube
Watch on YouTube
Watch on YouTube

o3 Mini FAQ

o3 Mini: capabilities, use cases, limits, and TokenHub guidance.

How should teams view o3 Mini?+

o3 Mini is a OpenAI model for cost-efficient STEM and coding reasoning.

What is o3 Mini best for?+

Best for mathematical reasoning, scientific reasoning and code reasoning, especially when speed and cost efficiency is the priority.

What is o3 Mini's main strength?+

Key strength: strong STEM and coding reasoning in a smaller text-only model.

Is o3 Mini always the best choice?+

It is deprecated or superseded, so it is a poor default for new integrations. Use GPT-5.4 Mini for new integrations.

What is the safest setup?+

Confirm availability; use the current recommended model for new integrations.

Ready to use o3-mini?

Use one API key to access o3-mini and more AI models through TokenHub.

Create API key