Home / Model Explorer / DeepSeek V4 Flash

DeepSeek V4 Flash API

DeepSeek·1M contextChatReasoningTool callingVision

V4.1 Flash Compatible Name · Ultra-Fast Multimodal

Overview

About this model

Compatible model name for DeepSeek V4.1 Flash. Officially, the old deepseek-v4-flash has been routed to V4.1 Flash, supporting 1M context, native visual understanding, tool calling, and thinking mode.

Capabilities

Capability matrix

TextSupported
VisionSupported
Audio—
ToolsSupported
JSONSupported
ReasoningSupported

Pricing and groups

Routes and pricing

DeepSeek Native Official90% of baseOperational

deepseek Full-power version, high concurrency

Input

$0.1342 / 1M

Output

$0.5367 / 1M

Cache read

$0.0027 / 1M

Cache write

—

Recent availabilityPast 24 hours · Active probes100%
24h ago16h ago8h agoNow

Prices in USD per 1M tokens. Rates follow provider adjustments; live prices in Model Explorer as the reference.

Snapshot: 2026-10-07 23:09. Availability reflects active probes over the past 24 hours; performance reflects actual calls.

API

Integration example

All models share one OpenAI-compatible endpoint. To switch models, change the model field. One key accesses every model.

Python · OpenAI SDK
# pip install openai
from openai import OpenAI

client = OpenAI(
    base_url="https://api.yomiapi.com/v1",
    api_key="sk-your-yomi-key",
)

resp = client.chat.completions.create(
    model="deepseek-v4-flash",
    messages=[{"role": "user", "content": "Hello"}],
)
print(resp.choices[0].message.content)
curl · OpenAI compatible
curl https://api.yomiapi.com/v1/chat/completions \
  -H "Authorization: Bearer sk-your-yomi-key" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "deepseek-v4-flash",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

Connect to DeepSeek V4 Flash now

Receive credits on registration, access every model with one key, and connect through low-latency global nodes.

Get a free API key