Grok 4.20 Non-Reasoning

Model ID: grok-4.20-0309-non-reasoning · Type: chat · Provider: xAI

Endpoints: /v1/chat/completions · /v1/messages

Pricing

Input (per 1M tokens)$1.25 USD
Output (per 1M tokens)$2.5 USD
Cache read (per 1M tokens)$0.2 USD
from openai import OpenAI

client = OpenAI(api_key="sk-...", base_url="https://api.router.ai/v1")
resp = client.chat.completions.create(
    model="grok-4.20-0309-non-reasoning",
    messages=[{"role": "user", "content": "Hello"}],
)
print(resp.choices[0].message.content)

FAQ

How much does Grok 4.20 Non-Reasoning cost on AI中转站?

Grok 4.20 Non-Reasoning (`grok-4.20-0309-non-reasoning`) is billed per usage at $1.25/1M in · $2.5/1M out, in USD. Current pricing is always listed at https://ai-zzz.com/models/grok-4.20-0309-non-reasoning.

How do I call Grok 4.20 Non-Reasoning through AI中转站?

Send a request to https://api.router.ai/v1/v1/chat/completions with the header `Authorization: Bearer <your API key>` and `"model": "grok-4.20-0309-non-reasoning"`. The API is OpenAI-compatible, so any OpenAI SDK works by changing base_url to https://api.router.ai/v1 — no other code change.

Which endpoints does Grok 4.20 Non-Reasoning support?

Grok 4.20 Non-Reasoning can be called on: /v1/chat/completions; /v1/messages.

What is Grok 4.20 Non-Reasoning's context window?

Grok 4.20 Non-Reasoning accepts up to 1,000,000 input tokens and can return up to 32,768 output tokens. Requests exceeding the input limit are rejected before reaching the model.

What can Grok 4.20 Non-Reasoning do?

Grok 4.20 Non-Reasoning supports: vision, function_calling, prompt_caching.

Who makes Grok 4.20 Non-Reasoning?

Grok 4.20 Non-Reasoning is a chat model from xAI, available through the AI中转站 gateway with the same API key as every other model.

Call it through the AI中转站 OpenAI-compatible endpoint (API base: https://api.router.ai/v1). AI agents can discover and call every model on this gateway through MCP (https://mcp.router.ai/mcp) with no manual integration.

API reference · All models