qwen3.6-flash

Text

Qwen 3.6 lightweight tier with text, image, and video input. Low latency and low cost, well-suited for high-volume classification, extraction, summarization, and simple agent workflows.

Input price

$0.123830% off

Market: $0.1769

/ M

Output price

$0.743330% off

Market: $1.0619

/ M

Cache read

$0.012330% off

Market: $0.0176

/ M

Cache write

$0.154830% off

Market: $0.2211

/ M

Input

Text

Output

Text

Dynamic pricing

Prices vary by usage tier and request conditions

Tiered price table

  • baseLength ≤ 256K
    Input$0.1238
    Output$0.7433
    Cache read$0.0123
    Cache write$0.1548
  • tier_2
    Input$0.4129
    Output$2.9734
    Cache read$0.0495
    Cache write$0.6194
TierInputOutputCache readCache write
baseLength ≤ 256K
$0.1238$0.7433$0.0123$0.1548
tier_2
$0.4129$2.9734$0.0495$0.6194

API

Quick Start

Drop-in code to call this model with EasyAPI's OpenAI-compatible API.

1

Get your API key

Create an API key from your dashboard and set it as an environment variable:

Replace sk-xxxxxxxx with your API token from the console (Bearer token).

Create API token
shell
export EASYAPI_API_KEY=sk-xxxxxxxx
2

Make your first request

Use qwen/qwen3.6-flash with the EasyAPI API:

cURL
curl https://token.easyapi.com/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $EASYAPI_API_KEY" \
  -d '{"model":"qwen/qwen3.6-flash","messages":[{"role":"user","content":"Hello!"}]}'
3

Enable streaming

Add "stream": true to your request body to receive responses as server-sent events:

curl
curl -N https://token.easyapi.com/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $EASYAPI_API_KEY" \
  -d '{"model":"qwen/qwen3.6-flash","messages":[{"role":"user","content":"Hello!"}],"stream":true}'

Endpoint

Sends a chat completion request. Supports streaming and non-streaming modes.

POSThttps://token.easyapi.com/v1/chat/completions
Authorization
Bearer $EASYAPI_API_KEY
Content-Type
application/json
Model
qwen/qwen3.6-flash

发送对话补全请求,支持流式与非流式。 Docs

GEThttps://token.easyapi.com/v1/models
Authorization
Bearer $EASYAPI_API_KEY

获取当前账户可用模型列表。 Docs

Parameters

NameTypeDefaultDescription
modelstring模型 ID,如 openai/gpt-4o、anthropic/claude-3-5-sonnet
messagesarray对话消息列表,每项含 role 与 content
streambooleanfalse设为 true 时以 SSE 流式返回
temperaturefloat1采样温度,越高越随机
max_tokensinteger单次回复最大生成 Token 数
top_pfloat1核采样,限制候选词累积概率
frequency_penaltyfloat0频率惩罚,降低重复用词
presence_penaltyfloat0存在惩罚,鼓励新话题

Using third-party SDKs

Official OpenAI and other SDKs work by swapping the base URL. See integration guide