Back to Models
DeepSeek V4 Flash (Free)
Available
Compare all DeepSeek models — V4 Pro vs Flash vs V3.2 vs V3.1 vs V3-0324
Description
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with a 1M-token context window, built for fast inference, high-throughput workloads, reasoning, coding, and agent workflows.
At a Glance
Key pricing and model details available for this model.
Input price
Free
per 1M tokens
Output price
Free
per 1M tokens
Context window
42K
tokens
Hallucination rate
0%
Token Pricing
Token pricing normalized to per-million-token rates.
Input / 1M tokens
Free
Output / 1M tokens
Free
Cache Read / 1M tokens
Free
Token Pricing Details
Rates are shown per 1M tokens for easier comparison.
| Input / 1M tokens | Free |
| Input unit | 1M tokens |
| Output / 1M tokens | Free |
| Output unit | 1M tokens |
| Cache Read / 1M tokens | Free |
| Cache Read unit | 1M tokens |
Feature Availability
Capabilities explicitly listed in the current payload.
LLM
Available
Yes
Vision
Not listed
No
Function calling
Available
Yes
Reasoning
Available
Yes
Service tiers
Not listed
No
Supported Parameters
frequency_penalty
logit_bias
max_completion_tokens
presence_penalty
stop
temperature
tool_choice
tools
top_p
Code Samples
Quick start with the Routeway API
import OpenAI from 'openai';
const openai = new OpenAI({
baseURL: "https://api.routeway.ai/v1",
apiKey: "<YOUR_API_KEY>",
});
async function main() {
const completion = await openai.chat.completions.create({
model: "deepseek-v4-flash:free",
messages: [
{
role: "user",
content: "Explain quantum computing in simple terms"
}
]
});
console.log(completion.choices[0].message);
}
main();