Gemini 3.7 Flash
Description
Google's high-efficiency Gemini 3.7 Flash model optimized for agentic workflows, coding, complex reasoning, and long-context multimodal intelligence with tunable thinking.
At a Glance
Key pricing and model details available for this model.
Input price
$0.75
per 1M tokens
Output price
$3.75
per 1M tokens
Context window
1M
tokens
Hallucination rate
0%
Token Pricing
Token pricing normalized to per-million-token rates.
Service tier
Set service_tier in your API request to choose throughput and pricing.
Input / 1M tokens
$0.75
Output / 1M tokens
$3.75
Cache Read / 1M tokens
$0.07
Token Pricing Details
Rates are shown per 1M tokens for easier comparison. Pass service_tier in your request to use the selected tier.
| Input / 1M tokens | $0.75 |
| Input unit | 1M tokens |
| Output / 1M tokens | $3.75 |
| Output unit | 1M tokens |
| Cache Read / 1M tokens | $0.07 |
| Cache Read unit | 1M tokens |
Feature Availability
Capabilities explicitly listed in the current payload.
LLM
Available
Vision
Available
Function calling
Available
Reasoning
Available
Service tiers
Available
Service Tiers
Available via the service_tier parameter.
Reasoning Effort Levels
Supported Parameters
Code Samples
Quick start with the Routeway API
import OpenAI from 'openai';
const openai = new OpenAI({
baseURL: "https://api.routeway.ai/v1",
apiKey: "<YOUR_API_KEY>",
});
async function main() {
const completion = await openai.chat.completions.create({
model: "gemini-3.7-flash",
messages: [
{
role: "user",
content: "Explain quantum computing in simple terms"
}
]
});
console.log(completion.choices[0].message);
}
main();