Back to Models

Qwen3.8 Flash

qwenqwen/qwen3.8-flash

Alibaba Qwen's 3.8-generation Flash tier: the cost/latency entry point into the 3.8 line, with a 1M-token context window.

Context Length
1M
Max Output
131K
Input Priceper 1M tokens
$0.162/ 1M tokens
Output Priceper 1M tokens
$0.5076/ 1M tokens

Modalities

text→text

Pricing Breakdown

TypeRate
Input$0.162 / 1M tokens
Output$0.5076 / 1M tokens

Supported Parameters

temperaturemax_tokenstop_ptoolstool_choiceresponse_formatstop

API Usage Examples

Use the global api.therouter.ai endpoint shown below for new integrations; the legacy China accelerated endpoint is retired.

cURL
curl https://api.therouter.ai/v1/chat/completions   -H "Content-Type: application/json"   -H "Authorization: Bearer $THE_ROUTER_API_KEY"   -d '{
    "model": "qwen/qwen3.8-flash",
    "messages": [
      {"role": "user", "content": "Summarize the key points from this input."}
    ]
  }'
Customer Support