返回模型列表

Qwen Plus

qwenqwen/qwen-plus

Qwen3 中档主力 — 支持 thinking / 非 thinking 双模式,1M 上下文。

供应商自营与 TheRouter 服务形态的差异

供应商自己运营时

阿里云在百炼中将 qwen-plus 系列记录为 1M 上下文,并支持思考模式、Function Calling、内置工具与结构化输出。上游 OpenAI 兼容 thinking 示例通过 extra_body 使用非标准 enable_thinking 标志。

在 TheRouter 上

TheRouter 在 OpenAI 兼容 chat 路由上以 qwen/qwen-plus 提供该模型,目录上下文为 1,048,576、最大输出 32,768,支持参数包括 tools、tool_choice、response_format、reasoning 与 stop。因此以阿里云 1M 与能力表作为上游权威,同时标明 TheRouter 的参数命名与目录限制。

qwen-plus 是 Qwen3 商用线的主力机型。提供两种工作模式 — 适合日常对话的快速模式(`enable_thinking: false`),以及面向更难问题、会在响应中输出 `reasoning_content` 的深度模式(`enable_thinking: true`) — 而价格只是旗舰的一小部分。

1M context window 让它成为长文档问答、多文件代码阅读与长 agent 对话的强默认值。TheRouter 按请求在 bailian-cn 与 bailian-sg 之间选择更便宜的一侧;`enable_thinking` 与 `reasoning_content` 在两端均原样透传。

thinking / 非 thinking
按请求切换 `enable_thinking`。为 true 时模型会在最终答案旁返回 `reasoning_content` 字段。
1M 上下文
Qwen 商用机型中最大的上下文窗口之一 — 适合整篇文档或多文件场景。
成本平衡
$0.50 输入 / $1.50 输出 per MTok。相对 qwen-max 输入便宜约 5 倍、输出便宜约 5.3 倍。
双区域路由
Selector 按请求选择 bailian-cn / bailian-sg 中更便宜的一侧 — 客户端无需任何区域逻辑。
适合使用
生产环境的 Qwen 默认机型:长文档问答、多步 agent、代码阅读,以及希望保留可选深度推理但不愿支付旗舰价格的对话场景。
不适合使用
需要 Qwen 系列最强推理或编码时升级到 qwen3-max / qwen3-coder-plus;高吞吐低成本对话场景请改用 qwen-turbo 或 qwen-flash。
计费方式:$0.50 输入 / $1.50 输出 per MTok。TheRouter 按请求路由到 bailian-cn / bailian-sg 中更便宜的一侧。Thinking 模式不改变价格。
上下文长度
1.0M
最大输出
33K
输入价格每百万 tokens
$0.432每百万 Tokens
输出价格每百万 tokens
$1.30每百万 Tokens

模态能力

文本→文本

价格明细

类型费率
输入$0.432 每百万 Tokens
输出$1.30 每百万 Tokens

支持参数

temperaturemax_tokenstop_ptoolstool_choiceresponse_formatreasoningstop

模型规格

上下文长度TheRouter 目录为 1,048,576 token;阿里云将 qwen-plus 系列上下文标注为 1Mhelp.aliyun.com ↗已核实
最大输出TheRouter 目录为 32,768 token已核实
思考模式上游 qwen-plus 系列支持;OpenAI 兼容调用在上游使用 enable_thinking,TheRouter 暴露 reasoning 参数help.aliyun.com ↗已核实
Function Calling 与工具阿里云文本生成指南说明所有通用模型支持 Function Calling,并在 Plus 档 Qwen 推荐中列出内置工具help.aliyun.com ↗已核实
结构化输出阿里云能力表中 qwen-plus 系列路由支持;TheRouter 通过 response_format 暴露help.aliyun.com ↗已核实
模态文本输入 → 文本输出help.aliyun.com ↗已核实

基准成绩

BenchmarkDistributionScoreSource
Public Plus-tier benchmark table
阿里云文本生成指南推荐 Plus 档 Qwen 作为能力与成本均衡选择,但已获取的公开页面未为 `qwen-plus` 别名发布数值基准表。
—未公开披露—
Thinking-mode benchmark
深度思考文档说明如何开启 thinking 并列出 qwen-plus 系列支持,但未发布 qwen-plus 推理分数。
—未公开披露—
TheRouter latency / throughput benchmark
尚无针对 TheRouter `qwen/qwen-plus` 路由的公开 TTFT 或吞吐测量。设定生产 SLO 前应分别测量 thinking 开启和关闭状态。
—未公开披露—

API 使用示例

所有新集成都应使用下方示例中的全球端点 api.therouter.ai;旧中国加速端点已下线。

cURL
curl https://api.therouter.ai/v1/chat/completions   -H "Content-Type: application/json"   -H "Authorization: Bearer $THE_ROUTER_API_KEY"   -d '{
    "model": "qwen/qwen-plus",
    "messages": [
      {"role": "user", "content": "Summarize the key points from this input."}
    ]
  }'

API 使用指南

完整 API 参考 →

Chat 调用

默认非 thinking 路径使用 OpenAI 兼容 chat 端点。只有在确认 SDK 如何序列化厂商特定选项后,再添加 reasoning。

cURL
curl https://api.therouter.ai/v1/chat/completions \
  -H "Authorization: Bearer $THEROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"qwen/qwen-plus","messages":[{"role":"user","content":"Summarize this customer feedback in three bullets."}],"max_tokens":600}'

qwen 其他模型

同类模型

跨供应商的相似能力档位

动态与变更

2026-08-02

阿里云文本生成指南将 Plus 档 Qwen 定位为均衡默认选择

当前百炼指南推荐 Plus 档 Qwen 用于办公、聊天机器人、内容生成、文档处理,以及需要 1M 上下文、工具和成本均衡的 agent 或编程工作流。无后缀 qwen-plus 路由仍作为旧版 Qwen 别名列出,并具备同类能力列。

TheRouter 编辑重写help.aliyun.com ↗

常见问题

TheRouter 中的 qwen/qwen-plus 是什么?

它是 TheRouter 面向阿里云百炼商业 qwen-plus 别名提供的 OpenAI 兼容路由:文入文出、百万级上下文、32K 最大输出,并在目录中支持 function calling、结构化输出和可选 reasoning。

qwen-plus 支持 thinking mode 吗?

支持。阿里云深度思考文档将 qwen-plus 系列列为混合思考路由。上游示例使用 enable_thinking;TheRouter 暴露 reasoning,因此在将推理轨迹视为稳定契约前,应测试 SDK 实际发送的请求体。

什么时候选择 qwen-plus 而不是 qwen-long?

普通生产 chat、agent、结构化输出,以及可放入 1M token 内的文档任务应选择 qwen-plus。只有当上下文大小是瓶颈,并且可以接受更窄的长文档路由时才选择 qwen-long。

qwen-plus 是固定的 Qwen3.7 checkpoint 吗?

不是。公开路由是无后缀商业别名。阿里云另行列出带日期的 Plus 快照(如 Qwen3.7 Plus)和旧 qwen-plus 快照;因此版本特定的回归或基准声明应在可用时绑定到快照路由。

事实档案 — 本页每条断言可在此回溯来源
来源URL采集于
上下文长度help.aliyun.com ↗2026-08-02已核实
最大输出——已核实
思考模式help.aliyun.com ↗2026-08-02已核实
Function Calling 与工具help.aliyun.com ↗2026-08-02已核实
结构化输出help.aliyun.com ↗2026-08-02已核实
模态help.aliyun.com ↗2026-08-02已核实
Public Plus-tier benchmark tablehelp.aliyun.com ↗2026-08-02未知
Thinking-mode benchmarkhelp.aliyun.com ↗2026-08-02未知
TheRouter latency / throughput benchmarkhelp.aliyun.com ↗2026-08-02未知
阿里云文本生成指南将 Plus 档 Qwen 定位为均衡默认选择help.aliyun.com ↗2026-08-02已核实
TheRouter 中的 qwen/qwen-plus 是什么?help.aliyun.com ↗2026-08-02待核实
qwen-plus 支持 thinking mode 吗?help.aliyun.com ↗2026-08-02待核实
什么时候选择 qwen-plus 而不是 qwen-long?help.aliyun.com ↗2026-08-02待核实
qwen-plus 是固定的 Qwen3.7 checkpoint 吗?help.aliyun.com ↗2026-08-02待核实
帮助与联系