Home / Model Explorer / Grok 4.6

Grok 4.6 API

xAI·1M contextChatReasoningTool callingVision

xAI Latest flagship · 1M context

Overview

About this model

xAI Latest flagship reasoning model: stronger deep reasoning and tool calling, native 1M long context. Note: when the context exceeds 200k tokens, the entire request (input, output, and cache) is billed at double the rate.

Capabilities

Capability matrix

TextSupported
VisionSupported
Audio—
ToolsSupported
JSONSupported
ReasoningSupported

Pricing and groups

Routes and pricing

Grok route16% of baseDegraded

Support grok4.6 (context >200k generates long-context double billing, please control the context reasonably)

Input

$0.32 / 1M

Output

$0.96 / 1M

Cache read

$0.08 / 1M

Cache write

—

Recent availabilityPast 24 hours · Active probes94.4%
24h ago16h ago8h agoNow

Prices in USD per 1M tokens. Rates follow provider adjustments; live prices in Model Explorer as the reference.

Snapshot: 2026-10-07 23:09. Availability reflects active probes over the past 24 hours; performance reflects actual calls.

API

Integration example

All models share one OpenAI-compatible endpoint. To switch models, change the model field. One key accesses every model.

Python · OpenAI SDK
# pip install openai
from openai import OpenAI

client = OpenAI(
    base_url="https://api.yomiapi.com/v1",
    api_key="sk-your-yomi-key",
)

resp = client.chat.completions.create(
    model="grok-4.6",
    messages=[{"role": "user", "content": "Hello"}],
)
print(resp.choices[0].message.content)
curl · OpenAI compatible
curl https://api.yomiapi.com/v1/chat/completions \
  -H "Authorization: Bearer sk-your-yomi-key" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "grok-4.6",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

Connect to Grok 4.6 now

Receive credits on registration, access every model with one key, and connect through low-latency global nodes.

Get a free API key