Qwen3.6-Flash

通义千问 3.6 快速预览, 低价高速, 含显式 cache.

Model ID: qwen3.6-flash · Type: chat · Provider: Alibaba

Endpoints: /v1/chat/completions · /v1/messages

Pricing

Input (per 1M tokens)$0.175 USD
Output (per 1M tokens)$1.05 USD
Cache read (per 1M tokens)$0.0175 USD
Cache write 5m (per 1M tokens)$0.21875 USD

Context tiers

≤ 256000 tokens$0.175 / 1M in$1.05 / 1M out
> 256000$0.7 / 1M in$2.8 / 1M out
from openai import OpenAI

client = OpenAI(api_key="sk-...", base_url="https://api.router.ai/v1")
resp = client.chat.completions.create(
    model="qwen3.6-flash",
    messages=[{"role": "user", "content": "Hello"}],
)
print(resp.choices[0].message.content)

FAQ

How much does Qwen3.6-Flash cost on 370.AI?

Qwen3.6-Flash (`qwen3.6-flash`) is billed per usage at $0.175/1M in · $1.05/1M out, in USD. Current pricing is always listed at https://www.370.ai/models/qwen3.6-flash.

How do I call Qwen3.6-Flash through 370.AI?

Send a request to https://api.router.ai/v1/v1/chat/completions with the header `Authorization: Bearer <your API key>` and `"model": "qwen3.6-flash"`. The API is OpenAI-compatible, so any OpenAI SDK works by changing base_url to https://api.router.ai/v1 — no other code change.

Which endpoints does Qwen3.6-Flash support?

Qwen3.6-Flash can be called on: /v1/chat/completions; /v1/messages.

Who makes Qwen3.6-Flash?

Qwen3.6-Flash is a chat model from Alibaba, available through the 370.AI gateway with the same API key as every other model.

Call it through the 370.AI OpenAI-compatible endpoint (API base: https://api.router.ai/v1). AI agents can discover and call every model on this gateway through MCP (https://mcp.router.ai/mcp) with no manual integration.

API reference · All models