GPT-5.4 Mini

gpt-5.4-mini

GPT-5.4 Mini is the smaller and faster member of the GPT-5.4 family. OpenAI documentation positions mini models for lower-latency workloads while retaining coding, tool use, multimodal reasoning, and strong instruction following. It is a good fit for well-scoped production tasks, subagents, and applications that need many fast calls.

Context Window

400K tokens

Maximum Output

128K tokens

Release Date

Mar 17, 2026

Modalities

GPT-5.4 Mini Pricing

这是价格描述

GPT-5.4 Mini API Capabilities

Reasoning

Supported

Tool calling

Supported

Temperature parameter

Not supported

Attachments

Supported

Knowledge Base

2025-08-31

Endpoint Protocols

Completions API

GPT-5.4 Mini Model Highlights

GPT-5.4 Mini brings configurable reasoning, coding and visual computer-work capabilities to faster, high-volume workloads.

Efficient High-Volume Reasoning

Retains GPT-5.4-class reasoning in a faster, more efficient model intended for large numbers of production requests.

Compact Coding Capability

Handles code generation, analysis and multi-file modifications while targeting lower-latency development workflows.

Visual Computer Reasoning

Understands screenshots and interface state to reason about actions in browser and desktop environments.

GPT-5.4 Mini Use Cases

GPT-5.4 Mini fits high-volume coding assistance, visual computer automation and delegated Agent tasks that benefit from faster execution.

High-Volume Code Assistance

Handles frequent code-generation, explanation and review requests across developer tools while maintaining practical response times.

Interface Automation

Interprets interface screenshots, follows application state and completes repeatable multi-step browser or desktop tasks.

Delegated Agent Subtasks

Executes bounded research, coding or analysis assignments within a larger workflow and returns structured results to the coordinating Agent.

How to Use GPT-5.4 Mini via the TokenHub API

Create API key
import OpenAI from "openai"

const client = new OpenAI({
  apiKey: process.env.TOKENHUB_API_KEY,
  baseURL: "https://us-api.tokenhub.com/v1",
})

const result = await client.chat.completions.create({})
console.log(result.choices[0]?.message?.content)

GPT-5.4 Mini Benchmarks

Index score
Artificial Analysis Intelligence IndexArtificial Analysis broad capability aggregate40
Artificial Analysis Coding IndexArtificial Analysis software task aggregate51.5
Knowledge & Reasoning
GPQAAdvanced science problem solving87.5%
HLEBroad expert-level exam set26.6%
Coding & Engineering
SciCodeScientific coding challenges49.9%
Terminal-Bench HardHard terminal task execution52.3%
Instruction Following & Agent Tasks
IFBenchPrompt constraint adherence73.3%
AA-LCRLong-context reasoning69.3%
τ²-BenchAgent workflow tasks83.3%

Metrics sourced from Artificial Analysis

Media and Discussions

Selected public videos and posts related to this model.

X (Twitter)

View post on X
View post on X
View post on X

Reddit

YouTube

Watch on YouTube
Watch on YouTube
Watch on YouTube

GPT-5.4 Mini FAQ

GPT-5.4 Mini: capabilities, use cases, limits, and TokenHub guidance.

How is GPT-5.4 Mini positioned?+

GPT-5.4 Mini is a OpenAI model for fast, efficient coding, tool use, and multimodal work.

Where does GPT-5.4 Mini add value?+

Best for routine coding assistance, tool-heavy automation and image and video understanding, especially when speed and cost efficiency is the priority.

What is GPT-5.4 Mini's practical edge?+

Key strength: strong coding and tool use with substantially lower latency and cost.

Which constraint matters most?+

It trades some peak quality for better speed or cost. For maximum answer quality, consider GPT-5.4.

How do I integrate GPT-5.4 Mini safely?+

Use the exact ID shown by TokenHub; follow your account docs and verify current features.

Ready to use GPT-5.4 Mini?

Use one API key to access GPT-5.4 Mini and more AI models through TokenHub.

Create API key