GPT-4o Mini

gpt-4o-mini

GPT-4o Mini is the fast and affordable small model in the GPT-4o family. OpenAI docs position it for focused tasks with text and image input, structured outputs, fine-tuning, and distillation workflows. It is best introduced as a lightweight multimodal production model rather than a reduced copy of GPT-4o.

Context Window

128K tokens

Maximum Output

16.4K tokens

Release Date

Jul 18, 2024

Modalities

GPT-4o Mini Pricing

Input PriceOutput PriceCache Read
$0.15/M$0.6/M$0.075/M

GPT-4o Mini API Capabilities

Reasoning

Not supported

Tool calling

Supported

Temperature parameter

Supported

Attachments

Supported

Knowledge Base

2023-09-01

Endpoint Protocols

Completions APIMessages APIgemini

GPT-4o Mini Model Highlights

GPT-4o Mini provides fast, economical text and image understanding for focused tasks and high-volume applications.

Fast Focused Processing

Handles well-defined tasks with response speed and processing economy suited to large numbers of requests.

Compact Vision

Accepts text and image inputs to interpret screenshots, photographs and document pages in a smaller model.

Task Specialization

Can be adapted to focused application behavior, making it useful for repeatable tasks with stable requirements.

GPT-4o Mini Use Cases

GPT-4o Mini fits high-volume intent classification, multilingual text processing and economical visual information extraction.

Intent Classification

Assigns user messages to predefined intents for support routing, workflow selection or analytics.

Multilingual Content Processing

Translates short content, extracts search terms and generates consistent tags across supported languages.

Visual Field Extraction

Extracts requested fields from forms, receipts or document images and returns concise text records.

How to Use GPT-4o Mini via the TokenHub API

Create API key
import OpenAI from "openai"

const client = new OpenAI({
  apiKey: process.env.TOKENHUB_API_KEY,
  baseURL: "https://us-api.tokenhub.com/v1",
})

const result = await client.chat.completions.create({})
console.log(result.choices[0]?.message?.content)

GPT-4o Mini Benchmarks

GPT-4o mini

Index score
Artificial Analysis Intelligence IndexArtificial Analysis broad capability aggregate6.9
Artificial Analysis Math IndexArtificial Analysis math reasoning aggregate14.7
Knowledge & Reasoning
MMLU-ProAdvanced multi-task knowledge64.8%
GPQAAdvanced science problem solving42.6%
HLEBroad expert-level exam set4%
Coding & Engineering
LiveCodeBenchLive coding problems23.4%
SciCodeScientific coding challenges22.9%
Math
MATH-500Advanced math problem solving78.9%
AIMECompetition math problems11.7%
AIME 2025Competition math problems14.7%
Instruction Following & Agent Tasks
IFBenchPrompt constraint adherence31.0%

Metrics sourced from Artificial Analysis

Media and Discussions

Selected public videos and posts related to this model.

X (Twitter)

View post on X
View post on X
View post on X

Reddit

YouTube

Watch on YouTube
Watch on YouTube
Watch on YouTube

Frequently asked questions about GPT-4o Mini

Understand what GPT-4o Mini is, its best uses, distinguishing strengths, practical tradeoffs, and safe TokenHub integration guidance.

Where does GPT-4o Mini sit within its provider’s model family?+

GPT-4o Mini is a compact omni model designed for fast, economical text and image workloads. It has been retired from ChatGPT, while API availability may remain; check TokenHub’s current listing.

Which production scenarios suit GPT-4o Mini?+

Best-fit scenarios include customer-support automation, large-scale classification and routing, and analysis of text and visual inputs. Test representative inputs and define measurable acceptance criteria before production.

What makes GPT-4o Mini stand out for large-scale classification and routing?+

Key strengths include fast response times, cost-efficient scaling, and combined text and image understanding. This combination is especially useful for large-scale classification and routing.

What tradeoffs should developers consider with GPT-4o Mini?+

Consider another model when the task requires the provider’s strongest reasoning capability, quality matters more than speed or cost, or the workflow cannot include human review for important decisions. Verify important factual, legal, financial, medical, or operational outputs with qualified human review.

How can a team safely start using GPT-4o Mini on TokenHub?+

In TokenHub, select the exact model identifier displayed for GPT-4o Mini, use the endpoint documented for your account, and authenticate with your TokenHub credentials. Check the current TokenHub documentation for supported text and image inputs, because platform exposure can differ from the provider’s full model capabilities.

Ready to use GPT-4o Mini?

Use one API key to access GPT-4o Mini and more AI models through TokenHub.

Create API key