HomeVastLLM Catalog › DeepSeek V4.1 Flash

DeepSeek V4.1 Flash

by DeepSeek · Language Models

Context 1049K tokensVisionReasoning / thinkingFunction calling

Overview

DeepSeek-V4.1-Flash is the smallest member of DeepSeek’s new model architecture family and natively supports multimodal visual understanding. The architecture targets stronger capabilities, faster inference and greater throughput, with room to scale to larger model sizes.

Context window1049K tokens
Max output393K tokens
Input → Outputtext, image → text
CapabilitiesVision, Reasoning / thinking, Function calling

Pricing

$0.1943 in · $0.777 out / 1M

UsageRate (USD)Unit
Input tokens$0.19per 1M tokens
Cached input (cache hit)$0per 1M tokens
Output tokens$0.78per 1M tokens

TuneRouter live platform pricing · currency: USD · pre-deduction settled by actual usage

Call it in 30 seconds

OpenAI-compatible · base URL https://api.tunerouter.com/v1/

POST/chat/completions
curl https://api.tunerouter.com/v1/chat/completions \
  -H "Authorization: Bearer $TUNEROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model": "deepseek-v4.1-flash", "messages": [{"role": "user", "content": "Hello!"}]}'

Model ID: deepseek-v4.1-flash · works with OpenAI SDKs — just point base_url at api.tunerouter.com

More from DeepSeek

← Back to full catalog (101 models)