Gemini 3.6 Flash
Total Context
1M
Max Output
65.5K
Released
N/A
gemini-3.7-flashGemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step problem solving.
Context Window
1M tokens
Maximum Output
65.5K tokens
Release Date
Aug 13, 2026
Modalities
| Input Price | Output Price | Cache Read |
|---|---|---|
| $0.375/M | $1.875/M | $0.0375/M |
Reasoning
Tool calling
Temperature parameter
Attachments
Knowledge Base
Endpoint Protocols
Gemini 3.7 Flash combines native multimodal reasoning, adjustable thinking, and a one-million-token input context for broad analysis workloads.
The model reasons across text, images, video, audio, and PDF inputs while producing text responses.
Low, medium, and high thinking levels let developers align reasoning depth with workload complexity.
An input limit of 1,048,576 tokens supports extensive documents and large collections of multimodal material in one request.
Gemini 3.7 Flash is suited to multimedia analysis, long-document synthesis, and reasoning over mixed-format technical information.
Analyze images, recorded audio, and video together with text to identify events, themes, and supporting evidence.
Process extensive PDFs and text collections, trace related details, and produce organized summaries or extracted findings.
Compare specifications, diagrams, logs, and explanatory text to answer complex questions and document the reasoning.
import OpenAI from "openai"
const client = new OpenAI({
apiKey: process.env.TOKENHUB_API_KEY,
baseURL: "https://us-api.tokenhub.com/v1",
})
const result = await client.chat.completions.create({})
console.log(result.choices[0]?.message?.content)| Index score | ||
|---|---|---|
| Artificial Analysis Intelligence Index | Artificial Analysis broad capability aggregate | 56 |
| Artificial Analysis Coding Index | Artificial Analysis software task aggregate | 76.1 |
Metrics sourced from Artificial Analysis
Common questions about using Gemini 3.7 Flash on TokenHub.
Gemini 3.7 Flash is Google’s Flash-series workhorse model for coding, agents, web development, and complex knowledge work.
It is well suited to software engineering, debugging, multi-step agent workflows, UI and web generation, and document-heavy knowledge tasks.
Google highlights stronger instruction following, multi-step planning, tool use, first-pass code quality, and design adherence than Gemini 3.6 Flash.
Review important outputs and test representative prompts. Choose another model when its quality, latency, supported features, or TokenHub price better matches your workload.
Select the model in a TokenHub API request and use your account endpoint and credentials. Google lists it for the Gemini API and AI Studio; Google’s introductory API price is $0.75 per million input tokens and $3.75 per million output tokens through 2026.
Use one API key to access Gemini 3.7 Flash and more AI models through TokenHub.
Gemini 3.7 Flash Media and Demos
Verified public posts, videos, and discussions about Gemini 3.7 Flash.
X (Twitter)
Reddit
YouTube