通义千问 3.7 Flash 原生视觉语言模型: 多模态理解与 Agent 执行能力全面升级, 万物识别/空间智能增强, 多模态 Coding 优化, 1M 上下文 + 推理
Model ID: qwen3.7-flash · Type: chat · Provider: Alibaba
Endpoints: /v1/chat/completions · /v1/messages
| Input (per 1M tokens) | $0.021 USD |
| Output (per 1M tokens) | $0.091 USD |
| Cache read (per 1M tokens) | $0.0021 USD |
| Cache write 5m (per 1M tokens) | $0.0266 USD |
| ≤ 32000 tokens | $0.021 / 1M in | $0.091 / 1M out |
| ≤ 256000 tokens | $0.07 / 1M in | $0.28 / 1M out |
| > 256000 | $0.14 / 1M in | $0.56 / 1M out |
from openai import OpenAI
client = OpenAI(api_key="sk-...", base_url="https://api.router.ai/v1")
resp = client.chat.completions.create(
model="qwen3.7-flash",
messages=[{"role": "user", "content": "Hello"}],
)
print(resp.choices[0].message.content)Qwen3.7-Flash (`qwen3.7-flash`) is billed per usage at $0.021/1M in · $0.091/1M out, in USD. Current pricing is always listed at https://ai-zzz.com/models/qwen3.7-flash.
Send a request to https://api.router.ai/v1/v1/chat/completions with the header `Authorization: Bearer <your API key>` and `"model": "qwen3.7-flash"`. The API is OpenAI-compatible, so any OpenAI SDK works by changing base_url to https://api.router.ai/v1 — no other code change.
Qwen3.7-Flash can be called on: /v1/chat/completions; /v1/messages.
Qwen3.7-Flash accepts up to 1,000,000 input tokens and can return up to 65,536 output tokens. Requests exceeding the input limit are rejected before reaching the model.
Qwen3.7-Flash supports: vision.
Qwen3.7-Flash is a chat model from Alibaba, available through the AI中转站 gateway with the same API key as every other model.
Call it through the AI中转站 OpenAI-compatible endpoint (API base: https://api.router.ai/v1). AI agents can discover and call every model on this gateway through MCP (https://mcp.router.ai/mcp) with no manual integration.