Back to Models
DeepSeek V4 Flash
Activedeepseek-v4-flashDeepSeek V4 Flash — fast OpenAI-compatible model with automatic context caching and an optional thinking mode (reasoning tokens counted inside completion tokens).
Technical Specifications
Context Length131K tokens
Max Output8K tokens
Available Providers1
Lowest Input Price$0.14/1M tokens
Lowest Output Price$0.28/1M tokens
Capabilities
chatcodefunction_calling
Provider Availability & Pricing
| Provider | Provider Model ID | Input Price (per 1M tokens) | Output Price (per 1M tokens) | Status |
|---|---|---|---|---|
DeepSeek | deepseek-v4-flash | $0.14 | $0.28 | active |
Code Examples
curl https://api.tokligence.ai/v1/chat/completions \
-H "Content-Type: application/json" \
-H "Authorization: Bearer YOUR_API_KEY" \
-d '{
"model": "deepseek-v4-flash",
"messages": [
{
"role": "user",
"content": "Hello, how are you?"
}
],
"max_tokens": 1024
}'Ready to use DeepSeek V4 Flash?
Try it out in our interactive playground or integrate it into your application with just a few lines of code.