Home / Model Explorer / Gemini 3.7 Flash
Gemini 3.7 Flash API
Latest Release · Ultra-Fast and Cost-Effective
Overview
About this model
Google The newly released Flash model, 1M ultra-long context, official promotional price until 2026/12/31.
Capabilities
Capability matrix
Pricing and groups
Routes and pricing
Third-Party Reverse-Engineered
Input
$0.105 / 1M
Output
$0.525 / 1M
Cache read
$0.0105 / 1M
Cache write
—
Google Vertex Direct Connection, balancing stability and response speed
Input
$0.27 / 1M
Output
$1.35 / 1M
Cache read
$0.027 / 1M
Cache write
—
Prices in USD per 1M tokens. Rates follow provider adjustments; live prices in Model Explorer as the reference.
Snapshot: 2026-10-07 23:09. Availability reflects active probes over the past 24 hours; performance reflects actual calls.
Compare
Group comparison
API
Integration example
All models share one OpenAI-compatible endpoint. To switch models, change the model field. One key accesses every model.
# pip install openai
from openai import OpenAI
client = OpenAI(
base_url="https://api.yomiapi.com/v1",
api_key="sk-your-yomi-key",
)
resp = client.chat.completions.create(
model="gemini-3.7-flash",
messages=[{"role": "user", "content": "Hello"}],
)
print(resp.choices[0].message.content)curl https://api.yomiapi.com/v1/chat/completions \
-H "Authorization: Bearer sk-your-yomi-key" \
-H "Content-Type: application/json" \
-d '{
"model": "gemini-3.7-flash",
"messages": [{"role": "user", "content": "Hello"}]
}'Connect to Gemini 3.7 Flash now
Receive credits on registration, access every model with one key, and connect through low-latency global nodes.
Get a free API key
Yomi API