DeepSeek's Official Agent Integrations Guide: The [1m] Context Specifier and Model Mapping Table Every AI Gateway Must Adopt
DeepSeek's official agent docs now cover 15 tools. The routing-critical details: a bracket context-suffix notation ([1m]) and a server-side model-mapping table that overrides gateway-level model selection.
Archive item produced with AI assistance from the cited source and published without individual review. Editor of record: Joe Werner.

DeepSeek has quietly expanded its official API documentation with a dedicated Agent Integrations section covering 15 tools — Claude Code, GitHub Copilot, GitHub Copilot CLI, Kilo Code, WorkBuddy/CodeBuddy, OpenCode, Oh My Pi, OpenClaw, AstrBot, Deep Code, Hermes, nanobot, Crush, Pi, Reasonix, and Langcli. For AI engineering teams, the surface-level story is ecosystem breadth. The routing story is something narrower and more urgent: an undocumented model specifier notation and an official model-mapping table that determines which DeepSeek tier your Claude requests actually land on.
What the official docs now specify
Two routing-critical facts are now spelled out for the first time in official DeepSeek API docs.
The [1m] context-length suffix. The recommended Claude Code configuration is:
export ANTHROPIC_MODEL=deepseek-v4-pro[1m]
export ANTHROPIC_DEFAULT_OPUS_MODEL=deepseek-v4-pro[1m]
export ANTHROPIC_DEFAULT_SONNET_MODEL=deepseek-v4-pro[1m]
export ANTHROPIC_DEFAULT_HAIKU_MODEL=deepseek-v4-flash
export CLAUDE_CODE_SUBAGENT_MODEL=deepseek-v4-flash
export CLAUDE_CODE_EFFORT_LEVEL=max
The [1m] suffix is a bracket notation that explicitly pins the 1M-token context window on the V4 Pro model. Without it, the API defaults to a shorter context tier. For agentic workloads — long codebase context, multi-file refactors, extended reasoning chains — the difference between default and [1m] is the difference between hitting context limits mid-session and completing the task.
The automatic model-mapping table. When Claude Code or the Claude Desktop app sends requests using Claude model names through the DeepSeek Anthropic API endpoint, DeepSeek automatically maps them:
| Claude model name prefix | Maps to |
|---|---|
claude-opus-* | deepseek-v4-pro |
claude-sonnet-* | deepseek-v4-flash |
claude-haiku-* | deepseek-v4-flash |
This mapping is applied server-side at https://api.deepseek.com/anthropic. It means that any team routing Claude Code through an AI gateway to DeepSeek's Anthropic endpoint does not control model selection at the model-name level — the upstream mapping table overrides it.
Why this matters for AI engineering teams
Context cost is now a routing dimension. The [1m] suffix is not just a convenience — it is an opt-in to the higher-cost 1M-context tier. DeepSeek's pricing structure distinguishes context window sizes; the [1m] flag requests the full-context variant. Teams that route all requests with this flag active will see per-token costs reflect the 1M-context pricing even for short exchanges. The correct pattern is to pin [1m] only on the Opus-tier model (for orchestrator-level tasks) and route Haiku-equivalent requests to deepseek-v4-flash without the suffix — exactly the config in the official docs.
The Anthropic endpoint model mapping table breaks gateway-level model pinning. If your AI gateway is translating Claude model names and forwarding them to https://api.deepseek.com/anthropic, you do not get DeepSeek model name granularity — the mapping table applies. Teams that need to pin deepseek-v4-pro explicitly must use the OpenAI-compatible endpoint at https://api.deepseek.com and set the ANTHROPIC_MODEL env var to the explicit DeepSeek model ID rather than routing through the Anthropic format. These are two separate routing surfaces with different behavior.
Web search adds a second billing event. The DeepSeek API natively supports Claude Code's web search tool. When a request triggers web search, the API issues a second LLM call to summarize retrieved content. This means web-search-capable agentic sessions generate two billing events per tool invocation. Teams building cost dashboards or per-session attribution need to account for this hidden second call in their usage accounting.
The ecosystem now includes 15 officially documented integrations. DeepSeek has moved beyond Claude Code and OpenCode to a much broader agent catalog. Tools like Kilo Code, Oh My Pi, AstrBot, Hermes, and OpenClaw all have dedicated per-tool docs. Each has its own model configuration pattern. Teams operating a gateway that fronts multiple agent tools should treat this as a signal that DeepSeek intends its Anthropic-compatible endpoint to be a primary backend for agent workloads — and plan their routing policies accordingly.
The router/operator angle
For teams running an AI gateway in front of DeepSeek's Anthropic-compatible endpoint, the key decision is whether to route at the gateway level or delegate model selection to the environment variable layer.
Gateway-level routing to DeepSeek: If your gateway translates Claude model names and forwards to the DeepSeek Anthropic endpoint, the server-side mapping table controls final model assignment. You lose the ability to distinguish V4 Pro vs V4 Flash at the gateway — the mapping is determined by the prefix of the requested Claude model name (opus → pro, sonnet/haiku → flash). This is fine if that coarse split aligns with your cost and quality targets.
Env-var-level pinning (bypassing gateway abstraction): The official config explicitly sets ANTHROPIC_MODEL=deepseek-v4-pro[1m], which overrides what Claude Code would normally select from the mapping table. This means the env var layer takes precedence over the Anthropic-format model name for the primary model slot — but sub-agent calls (via CLAUDE_CODE_SUBAGENT_MODEL) still route to flash separately.
Recommended routing policy for mixed-agent teams:
- Set the orchestrator model to
deepseek-v4-pro[1m]for Claude Code and any Opus-tier tool - Route Sonnet/Haiku-tier sub-agent calls to
deepseek-v4-flashwithout the[1m]suffix - Account for web search second-call billing in cost attribution
- Review each of the 15 official integration guides if you front multiple agent tools — each may have different model config patterns
What TheRouter users should watch
DeepSeek's Anthropic-compatible endpoint behavior — including the model mapping table and [1m] specifier — is relevant to any team routing Claude Code or other Anthropic-SDK tools through TheRouter to DeepSeek as an upstream. Verify that your provider routing configuration correctly handles the Anthropic endpoint format and that your cost accounting captures both primary and web-search-triggered billing events. The provider setup docs cover the base URL and auth configuration; the model-specifier details above are the operator-level config that must sit above the gateway layer.

DeepSeek Deprecates deepseek-chat Model Name: V4 Flash Migration and the New Anthropic API Routing Playbook
DeepSeek retires deepseek-chat and deepseek-reasoner on July 24, replacing them with explicit V4 model names and a new Anthropic API endpoint. Here is what changes for AI engineering teams routing through DeepSeek today.

DeepSeek Claude Code & OpenCode Integration: Official V4 API Setup Guide
DeepSeek's official integration guide for Claude Code, OpenCode, and OpenClaw — API setup, per-tier V4 model routing, and Anthropic-compatible endpoint configuration for coding agents.

Claude Code 2.1.275 Broke Every Gateway Proxy. 2.1.276 Fixed It the Same Day.
A new internal request tag in 2.1.275 caused 400 errors on every proxy-routed API call. 2.1.276 hotfixed it the same day. Breakdown of the failure, affected configs, and three secondary operator changes worth auditing.