Back to Models
Gemini 3.8 Flash
Available
Description
Google's Gemini 3.8 Flash is a fast multimodal model for coding, reasoning, and everyday agentic workloads.
At a Glance
Key pricing and model details available for this model.
Input price
$0.75
per 1M tokens
Output price
$3.75
per 1M tokens
Context window
1M
tokens
Hallucination rate
0%
Token Pricing
Token pricing normalized to per-million-token rates.
Input / 1M tokens
$0.75
Output / 1M tokens
$3.75
Cache Read / 1M tokens
$0.08
Cache Write (5m) / 1M tokens
$0.08
Token Pricing Details
Rates are shown per 1M tokens for easier comparison.
| Input / 1M tokens | $0.75 |
| Input unit | 1M tokens |
| Output / 1M tokens | $3.75 |
| Output unit | 1M tokens |
| Cache Read / 1M tokens | $0.08 |
| Cache Read unit | 1M tokens |
| Cache Write (5m) / 1M tokens | $0.08 |
Feature Availability
Capabilities explicitly listed in the current payload.
LLM
Available
Yes
Vision
Available
Yes
Function calling
Available
Yes
Reasoning
Available
Yes
Service tiers
Not listed
No
Reasoning Effort Levels
minimal
low
medium
high
Supported Parameters
frequency_penalty
logit_bias
max_completion_tokens
presence_penalty
reasoning_effort
stop
temperature
top_p
tool_choice
tools
Code Samples
Quick start with the Routeway API
import OpenAI from 'openai';
const openai = new OpenAI({
baseURL: "https://api.routeway.ai/v1",
apiKey: "<YOUR_API_KEY>",
});
async function main() {
const completion = await openai.chat.completions.create({
model: "gemini-3.8-flash",
messages: [
{
role: "user",
content: "Explain quantum computing in simple terms"
}
]
});
console.log(completion.choices[0].message);
}
main();