Home / Model Explorer / DeepSeek V4.1 Flash
DeepSeek V4.1 Flash API
Native multimodal · ultra-fast and cost-effective
Overview
About this model
DeepSeek V4.1 Flash, 552B MoE native multimodal model, 1M context, combining high speed, low cost, and strong Agent/reasoning capabilities; supports text and visual understanding, tool calling, and JSON output. The official API model name is deepseek-flash; the platform retains this compatibility name.
Capabilities
Capability matrix
Pricing and groups
Routes and pricing
deepseek Full-power version, high concurrency
Input
$0.1342 / 1M
Output
$0.5367 / 1M
Cache read
$0.0027 / 1M
Cache write
—
Prices in USD per 1M tokens. Rates follow provider adjustments; live prices in Model Explorer as the reference.
Snapshot: 2026-10-07 23:09. Availability reflects active probes over the past 24 hours; performance reflects actual calls.
API
Integration example
All models share one OpenAI-compatible endpoint. To switch models, change the model field. One key accesses every model.
# pip install openai
from openai import OpenAI
client = OpenAI(
base_url="https://api.yomiapi.com/v1",
api_key="sk-your-yomi-key",
)
resp = client.chat.completions.create(
model="deepseek-v4.1-flash",
messages=[{"role": "user", "content": "Hello"}],
)
print(resp.choices[0].message.content)curl https://api.yomiapi.com/v1/chat/completions \
-H "Authorization: Bearer sk-your-yomi-key" \
-H "Content-Type: application/json" \
-d '{
"model": "deepseek-v4.1-flash",
"messages": [{"role": "user", "content": "Hello"}]
}'Connect to DeepSeek V4.1 Flash now
Receive credits on registration, access every model with one key, and connect through low-latency global nodes.
Get a free API key
Yomi API