Anthropic
Claude Sonnet 5.5
Total Context
1M
Max Output
128K
Released
N/A
claude-haiku-5.5Claude Haiku 5.5 is Anthropic’s fast, efficient model for high-volume applications where responsiveness and cost matter. It is designed for summarization, classification, information extraction, request routing, and live customer support, making it well suited to everyday automation and interactive assistants. With a 1-million-token context window and adaptive thinking, Haiku 5.5 combines large-context processing with adjustable reasoning depth. It also works well as a specialized subagent, handling focused research, document lookups, and supporting coding tasks within larger workflows. Choose Haiku 5.5 for narrowly scoped work that needs to run quickly and economically at scale.
Context Window
1M tokens
Maximum Output
128K tokens
Release Date
Oct 7, 2026
Modalities
| Token Tier | Input Price | Output Price | Cache Read | Cache Create 5m |
|---|---|---|---|---|
| <=100K | $0.1/M | $0.5/M | $0.01/M | $0.125/M |
| >100K | $0.5/M | $2.5/M | $0.05/M | $0.625/M |
Reasoning
Tool calling
Temperature parameter
Attachments
Knowledge Base
Endpoint Protocols
Verified strengths and model characteristics.
Designed for frequent, latency-sensitive tasks, with lower pricing than Claude Haiku 4.5 and stronger coding and knowledge-work performance.
The first Haiku generation with adjustable effort lets developers balance reasoning depth and computational cost per task.
A 1M-token context window accepts text and images, with up to 128K output tokens in standard requests.
Practical workloads supported by official examples.
Classify incoming records, extract relevant fields, and route requests to the appropriate workflow.
Summarize documents and condense agent histories into concise context for subsequent work.
Delegate focused coding or document-reading subtasks while a larger model coordinates the overall project.
curl 'https://us-api.tokenhub.com/v1/messages' \
-X 'POST' \
-H "Authorization: Bearer $TOKENHUB_API_KEY"| Index score | ||
|---|---|---|
| Artificial Analysis Intelligence Index | Artificial Analysis broad capability aggregate | 37.8 |
Metrics sourced from Artificial Analysis
Model selection, limitations, and access.
Anthropic’s small model targets high-volume, latency-sensitive work, combining adaptive reasoning with text and image understanding.
Choose it for classification, extraction, summaries, live support, and narrow subagent tasks where response time and repeated-call cost matter.
Anthropic recommends larger models for complex agentic coding. Higher effort can increase cost and latency; prompts above 100K tokens also enter a higher pricing tier.
Use the identifier and endpoint documented by TokenHub after confirming access. Anthropic’s direct API ID is claude-haiku-5-5; do not assume it matches a TokenHub alias.
Anthropic lists $0.10 input/$0.50 output per million tokens for prompts up to 100K tokens, and $0.50/$2.50 above that threshold. TokenHub rates may differ.
Use one API key to access Claude Haiku 5.5 and more AI models through TokenHub.
Claude Haiku 5.5 Media and Demos
Verified public announcements, demonstrations, and community discussions.
X (Twitter)
Claude Haiku 5.5: Official Launch
Claude
View originalClaude Haiku 5.5: Workloads And Subagents
Claude
View originalClaude Haiku 5.5: Adjustable Reasoning Effort
Claude
View originalReddit
Claude Haiku 5.5: Official Reddit Announcement
ClaudeOfficial
View originalClaude Haiku 5.5: Community Coding Experiment
Escobar747
View originalClaude Haiku 5.5: Community Impressions
Outrageous-Exam9084
View originalYouTube
Claude Haiku 5.5: Model Review
Bijan Bowen
View originalClaude Haiku 5.5: Performance And Pricing Review
WorldofAI
View originalClaude Haiku 5.5: Strengths And Limitations
Universe of AI
View original