阿里云文本生成指南将 Plus 档 Qwen 定位为均衡默认选择
当前百炼指南推荐 Plus 档 Qwen 用于办公、聊天机器人、内容生成、文档处理,以及需要 1M 上下文、工具和成本均衡的 agent 或编程工作流。无后缀 qwen-plus 路由仍作为旧版 Qwen 别名列出,并具备同类能力列。
Qwen3 中档主力 — 支持 thinking / 非 thinking 双模式,1M 上下文。
阿里云在百炼中将 qwen-plus 系列记录为 1M 上下文,并支持思考模式、Function Calling、内置工具与结构化输出。上游 OpenAI 兼容 thinking 示例通过 extra_body 使用非标准 enable_thinking 标志。
TheRouter 在 OpenAI 兼容 chat 路由上以 qwen/qwen-plus 提供该模型,目录上下文为 1,048,576、最大输出 32,768,支持参数包括 tools、tool_choice、response_format、reasoning 与 stop。因此以阿里云 1M 与能力表作为上游权威,同时标明 TheRouter 的参数命名与目录限制。
qwen-plus 是 Qwen3 商用线的主力机型。提供两种工作模式 — 适合日常对话的快速模式(`enable_thinking: false`),以及面向更难问题、会在响应中输出 `reasoning_content` 的深度模式(`enable_thinking: true`) — 而价格只是旗舰的一小部分。
1M context window 让它成为长文档问答、多文件代码阅读与长 agent 对话的强默认值。TheRouter 按请求在 bailian-cn 与 bailian-sg 之间选择更便宜的一侧;`enable_thinking` 与 `reasoning_content` 在两端均原样透传。
| 类型 | 费率 |
|---|---|
| 输入 | $0.432 每百万 Tokens |
| 输出 | $1.30 每百万 Tokens |
| 上下文长度 | TheRouter 目录为 1,048,576 token;阿里云将 qwen-plus 系列上下文标注为 1Mhelp.aliyun.com ↗ | 已核实 |
| 最大输出 | TheRouter 目录为 32,768 token | 已核实 |
| 思考模式 | 上游 qwen-plus 系列支持;OpenAI 兼容调用在上游使用 enable_thinking,TheRouter 暴露 reasoning 参数help.aliyun.com ↗ | 已核实 |
| Function Calling 与工具 | 阿里云文本生成指南说明所有通用模型支持 Function Calling,并在 Plus 档 Qwen 推荐中列出内置工具help.aliyun.com ↗ | 已核实 |
| 结构化输出 | 阿里云能力表中 qwen-plus 系列路由支持;TheRouter 通过 response_format 暴露help.aliyun.com ↗ | 已核实 |
| 模态 | 文本输入 → 文本输出help.aliyun.com ↗ | 已核实 |
| Benchmark | Distribution | Score | Source |
|---|---|---|---|
Public Plus-tier benchmark table 阿里云文本生成指南推荐 Plus 档 Qwen 作为能力与成本均衡选择,但已获取的公开页面未为 `qwen-plus` 别名发布数值基准表。 | — | 未公开披露 | — |
Thinking-mode benchmark 深度思考文档说明如何开启 thinking 并列出 qwen-plus 系列支持,但未发布 qwen-plus 推理分数。 | — | 未公开披露 | — |
TheRouter latency / throughput benchmark 尚无针对 TheRouter `qwen/qwen-plus` 路由的公开 TTFT 或吞吐测量。设定生产 SLO 前应分别测量 thinking 开启和关闭状态。 | — | 未公开披露 | — |
所有新集成都应使用下方示例中的全球端点 api.therouter.ai;旧中国加速端点已下线。
curl https://api.therouter.ai/v1/chat/completions -H "Content-Type: application/json" -H "Authorization: Bearer $THE_ROUTER_API_KEY" -d '{
"model": "qwen/qwen-plus",
"messages": [
{"role": "user", "content": "Summarize the key points from this input."}
]
}'默认非 thinking 路径使用 OpenAI 兼容 chat 端点。只有在确认 SDK 如何序列化厂商特定选项后,再添加 reasoning。
curl https://api.therouter.ai/v1/chat/completions \
-H "Authorization: Bearer $THEROUTER_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"qwen/qwen-plus","messages":[{"role":"user","content":"Summarize this customer feedback in three bullets."}],"max_tokens":600}'当前百炼指南推荐 Plus 档 Qwen 用于办公、聊天机器人、内容生成、文档处理,以及需要 1M 上下文、工具和成本均衡的 agent 或编程工作流。无后缀 qwen-plus 路由仍作为旧版 Qwen 别名列出,并具备同类能力列。
它是 TheRouter 面向阿里云百炼商业 qwen-plus 别名提供的 OpenAI 兼容路由:文入文出、百万级上下文、32K 最大输出,并在目录中支持 function calling、结构化输出和可选 reasoning。
支持。阿里云深度思考文档将 qwen-plus 系列列为混合思考路由。上游示例使用 enable_thinking;TheRouter 暴露 reasoning,因此在将推理轨迹视为稳定契约前,应测试 SDK 实际发送的请求体。
普通生产 chat、agent、结构化输出,以及可放入 1M token 内的文档任务应选择 qwen-plus。只有当上下文大小是瓶颈,并且可以接受更窄的长文档路由时才选择 qwen-long。
不是。公开路由是无后缀商业别名。阿里云另行列出带日期的 Plus 快照(如 Qwen3.7 Plus)和旧 qwen-plus 快照;因此版本特定的回归或基准声明应在可用时绑定到快照路由。
| 来源 | URL | 采集于 | |
|---|---|---|---|
| 上下文长度 | help.aliyun.com ↗ | 2026-08-02 | 已核实 |
| 最大输出 | — | — | 已核实 |
| 思考模式 | help.aliyun.com ↗ | 2026-08-02 | 已核实 |
| Function Calling 与工具 | help.aliyun.com ↗ | 2026-08-02 | 已核实 |
| 结构化输出 | help.aliyun.com ↗ | 2026-08-02 | 已核实 |
| 模态 | help.aliyun.com ↗ | 2026-08-02 | 已核实 |
| Public Plus-tier benchmark table | help.aliyun.com ↗ | 2026-08-02 | 未知 |
| Thinking-mode benchmark | help.aliyun.com ↗ | 2026-08-02 | 未知 |
| TheRouter latency / throughput benchmark | help.aliyun.com ↗ | 2026-08-02 | 未知 |
| 阿里云文本生成指南将 Plus 档 Qwen 定位为均衡默认选择 | help.aliyun.com ↗ | 2026-08-02 | 已核实 |
| TheRouter 中的 qwen/qwen-plus 是什么? | help.aliyun.com ↗ | 2026-08-02 | 待核实 |
| qwen-plus 支持 thinking mode 吗? | help.aliyun.com ↗ | 2026-08-02 | 待核实 |
| 什么时候选择 qwen-plus 而不是 qwen-long? | help.aliyun.com ↗ | 2026-08-02 | 待核实 |
| qwen-plus 是固定的 Qwen3.7 checkpoint 吗? | help.aliyun.com ↗ | 2026-08-02 | 待核实 |