OpenAI
GPT-6.1 Sol
Total Context
1.1M
Max Output
128K
Released
N/A
gpt-5.4-nanoGPT-5.4 Nano is the lowest-cost, smallest GPT-5.4 option for high-volume simple tasks. OpenAI’s model guidance positions nano models around latency and cost efficiency rather than maximum reasoning depth. It should be used in descriptions for classification, routing, extraction, lightweight generation, and other predictable workflows.
Context Window
400K tokens
Maximum Output
128K tokens
Release Date
Mar 17, 2026
Modalities
| Input Price | Output Price | Cache Read |
|---|---|---|
| $0.2/M | $1.25/M | $0.02/M |
Reasoning
Tool calling
Temperature parameter
Attachments
Knowledge Base
Endpoint Protocols
GPT-5.4 nano prioritizes speed and processing economy for focused, high-volume text and image workloads.
Is optimized for simple, well-scoped tasks where rapid responses and economical execution are primary requirements.
Supports classification, extraction and ranking tasks that map large input sets into consistent labels, fields or ordered results.
Accepts up to 400,000 context tokens and image inputs, allowing focused processing across sizeable document collections.
GPT-5.4 nano is intended for large-scale classification, information extraction and ranking workloads with predictable output requirements.
Assigns predefined categories to high volumes of messages, documents or images for routing, analytics and moderation pipelines.
Extracts named fields from text, forms or document images and prepares consistent records for downstream systems.
Scores and orders search results, recommendations or support items according to explicit relevance criteria.
import OpenAI from "openai"
const client = new OpenAI({
apiKey: process.env.TOKENHUB_API_KEY,
baseURL: "https://us-api.tokenhub.com/v1",
})
const result = await client.chat.completions.create({})
console.log(result.choices[0]?.message?.content)| Index score | ||
|---|---|---|
| Artificial Analysis Intelligence Index | Artificial Analysis broad capability aggregate | 38.2 |
| Artificial Analysis Coding Index | Artificial Analysis software task aggregate | 43.9 |
| Knowledge & Reasoning | ||
| GPQA | Advanced science problem solving | 81.7% |
| HLE | Broad expert-level exam set | 26.5% |
| Coding & Engineering | ||
| SciCode | Scientific coding challenges | 46.9% |
| Terminal-Bench Hard | Hard terminal task execution | 42.4% |
| Instruction Following & Agent Tasks | ||
| IFBench | Prompt constraint adherence | 75.9% |
| AA-LCR | Long-context reasoning | 66% |
| τ²-Bench | Agent workflow tasks | 76.0% |
Metrics sourced from Artificial Analysis
GPT-5.4 Nano: capabilities, use cases, limits, and TokenHub guidance.
GPT-5.4 Nano is a OpenAI model for simple high-volume tasks at the lowest GPT-5.4-class cost.
Best for simple extraction and classification, structured data generation and high-volume requests, especially when low-cost scale is the priority.
Key strength: the lowest-cost GPT-5.4-class option for simple high-volume work.
It is optimized for simple work and is not the best choice for difficult reasoning or large engineering tasks. For maximum answer quality, consider GPT-5.4 Mini.
Use the exact ID shown by TokenHub; follow your account docs and verify current features.
Use one API key to access GPT-5.4 Nano and more AI models through TokenHub.
Media and Discussions
Selected public videos and posts related to this model.
X (Twitter)
Reddit
YouTube