Home / Model Explorer / Gemini 3.6 Flash

Gemini 3.6 Flash API

Google·1M contextChatTool callingVision

Limited-Time Offer · Ultra-Fast Response

Overview

About this model

Google Flash latest iteration of the series, 1M ultra-long context, official promotional price valid until 2026/12/31, a highly cost-effective choice.

Capabilities

Capability matrix

TextSupported
VisionSupported
Audio—
ToolsSupported
JSONSupported
Reasoning—

Pricing and groups

Routes and pricing

Gemini Official Direct Connection36% of baseOperational

Google Vertex Direct Connection, balancing stability and response speed

Input

$0.27 / 1M

Output

$1.35 / 1M

Cache read

$0.027 / 1M

Cache write

—

Recent availabilityPast 24 hours · Active probes99%
24h ago16h ago8h agoNow

Prices in USD per 1M tokens. Rates follow provider adjustments; live prices in Model Explorer as the reference.

Snapshot: 2026-10-07 23:09. Availability reflects active probes over the past 24 hours; performance reflects actual calls.

API

Integration example

All models share one OpenAI-compatible endpoint. To switch models, change the model field. One key accesses every model.

Python · OpenAI SDK
# pip install openai
from openai import OpenAI

client = OpenAI(
    base_url="https://api.yomiapi.com/v1",
    api_key="sk-your-yomi-key",
)

resp = client.chat.completions.create(
    model="gemini-3.6-flash",
    messages=[{"role": "user", "content": "Hello"}],
)
print(resp.choices[0].message.content)
curl · OpenAI compatible
curl https://api.yomiapi.com/v1/chat/completions \
  -H "Authorization: Bearer sk-your-yomi-key" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "gemini-3.6-flash",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

Connect to Gemini 3.6 Flash now

Receive credits on registration, access every model with one key, and connect through low-latency global nodes.

Get a free API key