Alibaba
Qwen3.8 Omni Flash
Total Context
1M
Max Output
131.1K
Released
N/A
qwen3.7-maxQwen3.7 Max is positioned as the flagship model of the Qwen3.7 generation, with official messaging centered on the “agent frontier.” Model cards emphasize agent-centric work, coding, office/productivity tasks, and long-horizon autonomous execution. It should be described as a broad high-capability model for complex work rather than merely a larger chat model.
Context Window
1M tokens
Maximum Output
65.5K tokens
Release Date
May 21, 2026
Modalities
| Input Price | Output Price | Cache Read | Cache Create 5m | Cache Read 5m |
|---|---|---|---|---|
| $1.7143/M | $5.1429/M | $0.3429/M | $2.1429/M | $0.1714/M |
Reasoning
Tool calling
Temperature parameter
Attachments
Knowledge Base
Endpoint Protocols
Qwen3.7 Max combines agent-oriented reasoning, strong coding and productivity capabilities, and a one-million-token context for complex text workloads.
The model is designed to plan and sustain multi-step work, making it useful when tasks require repeated decisions and long-term execution.
Its strengths cover programming, office work, and productivity tasks that require structured analysis and coordinated text generation.
A one-million-token context window supports large codebases, extensive documents, and long-running text interactions.
Qwen3.7 Max is suited to repository-scale development, long-running text agents, and analysis of extensive business or technical documents.
Examine dependencies across a large codebase, implement coordinated changes, and verify how modifications affect related components.
Maintain goals and intermediate results while progressing through multi-step research, coding, or operational workflows.
Review extensive text collections, connect evidence across sections, and produce structured summaries, comparisons, or action plans.
import OpenAI from "openai"
const client = new OpenAI({
apiKey: process.env.TOKENHUB_API_KEY,
baseURL: "https://us-api.tokenhub.com/v1",
})
const result = await client.chat.completions.create({})
console.log(result.choices[0]?.message?.content)Qwen3.7 Max
| Index score | ||
|---|---|---|
| Artificial Analysis Intelligence Index | Artificial Analysis broad capability aggregate | 46 |
| Artificial Analysis Coding Index | Artificial Analysis software task aggregate | 50.1 |
| Knowledge & Reasoning | ||
| GPQA | Advanced science problem solving | 92.3% |
| HLE | Broad expert-level exam set | 38.1% |
| Coding & Engineering | ||
| SciCode | Scientific coding challenges | 48.8% |
| Terminal-Bench Hard | Hard terminal task execution | 50.8% |
| Instruction Following & Agent Tasks | ||
| IFBench | Prompt constraint adherence | 80.5% |
| AA-LCR | Long-context reasoning | 69% |
| τ²-Bench | Agent workflow tasks | 94.7% |
Metrics sourced from Artificial Analysis
Qwen 3.7 Max: capabilities, use cases, limits, and TokenHub guidance.
Qwen 3.7 Max is a Alibaba Qwen model for flagship coding, reasoning, and agent workflows.
Best for complex coding, agent workflows and tool-heavy automation, especially when maximum answer quality is the priority.
Key strength: broad agent capabilities across coding and tool-driven tasks and hybrid thinking that can switch between deliberate and direct responses.
It uses more compute, so latency and cost can be higher. For multimodal input, consider Qwen 3.7 Plus.
Use the exact ID shown by TokenHub; follow your account docs and verify current features.
Use one API key to access Qwen3.7 Max and more AI models through TokenHub.
Media and Discussions
Selected public videos and posts related to this model.
X (Twitter)
Reddit
YouTube