IPIPCombo
Side-by-side comparisons

AI Assistants

DeepSeek vs Kimi vs Doubao vs Tongyi: which Chinese LLM should you actually use

Four Chinese LLMs have crossed the frontier line on coding and reasoning, with pricing a tenth of GPT-5 or Claude Opus. DeepSeek, Kimi, Doubao, and Tongyi differ on context window, ecosystem, and whether you can use them outside China. Here's how to pick.

Last verified: 2026-09-23
Dimension
DeepSeek (深度求索)
Open weights, strongest price-to-performance
Kimi (月之暗面)
256K context, agentic coding, strong Deep Research
Doubao (字节跳动)
Multimodal, adaptive thinking, low latency
Tongyi / Qwen (阿里巴巴)
Open ecosystem, 1M-token window on flagship models
Developer / companyDeepSeek AI (Hangzhou)Moonshot AI (北京)ByteDance / 火山引擎 (Volcano Engine)Alibaba Cloud / DashScope
Flagship model (Sep 2026)DeepSeek V3.2 ThinkKimi K3 / K2.7 CodeDoubao Seed 2.1 ProQwen 3.7-Plus / 3.8-Max
Context window128K (164K via third-party hosts)256K (1.05M on K3)256K (1M on weekly evolving)1M (256K on Qwen 3.7-Plus)
Coding performanceStrong (37% LiveCodeBench, 42% SWE-bench Verified)Strong on long-horizon agentic coding (K2.7 Code)Solid, good front-end and adaptive thinkingStrong, SOTA on SWE-bench among open weights (Qwen3-Coder)
Math / reasoningStrong (V3.2 Thinking mode for hard problems)Strong (K3 ranks high on public leaderboards)Solid, with selectable thinking lengthsStrong (Qwen 3.8-Max), strong on AIME-style tasks
Chinese language qualityExcellent — idiomatic, mainstream vocabularyExcellent — long-context Chinese synthesis is the headlineExcellent — strongest at colloquial / internet ChineseExcellent — strongest on formal / classical Chinese
Multimodal inputText only (image understanding is hosted-only)Image input (Kimi K2.6+)Image + video understanding on Seed 1.6+Image + video understanding (Qwen 3.7-Plus / Qwen3-VL)
API pricing (per 1M tokens)$0.26 input / $0.89 output (cache hits $0.07)$0.95 input / $4 output (cache hits $0.16)$0.25 input / $2 output (CN ¥6 / ¥30)$0.28–0.40 input / $1.10–1.60 output
Free consumer tierYes (deepseek.com web + app, low quotas)Yes — Adagio tier, 6 agent tasksYes (in 豆包 client), generous quotasYes (chat.qwen.ai + Qwen Work), generous
API availability outside ChinaYes — OpenAI-compatible at api.deepseek.comYes — api.moonshot.ai OpenAI-compatibleVia BytePlus ModelArk (international OpenAI-compatible)Yes — Singapore, Frankfurt, Virginia, Tokyo regions
Open weightsYes (V3 base, with MIT-ish terms)Yes (K2 series, modified MIT)No — closed weightsYes (Qwen3, Qwen3-Coder — Apache 2.0)
Best forCost-sensitive API workloads, self-hosted open-weight pipelinesAgentic coding on a budget, long-document Chinese researchMultimodal + low-latency assistants integrated with Douyin / LarkOpen ecosystem, full-stack on Alibaba Cloud, 1M-token window

Bottom line

If cost matters most and you can live with text-only, DeepSeek V3.2 Think is the rational default — open weights, frontier-class reasoning, and the cheapest cache-hit pricing of the four. If you need long-context Chinese research or agentic coding, Kimi K3 / K2.7 Code is the better pick at a 3-4× price premium. For multimodal and the most generous free consumer tier, Doubao Seed 2.1 Pro is hard to beat. For open ecosystems and 1M-token context on a flagship non-thinking model, Tongyi Qwen3-Plus / Max is the obvious choice. Most production setups stack two: DeepSeek for default cost-sensitive workloads, and one of Kimi / Qwen for the long-context or open-weights layer.

Full pages for each side

DeepSeek vs Kimi vs Doubao vs Tongyi: which Chinese LLM should you actually use