One key covers many models, but they are not interchangeable. The table below helps you choose by use case. Billing is unified in tokens - see Pricing.
| Model | Strength | Context | Modality | Best for |
|---|---|---|---|---|
DeepSeek | Reasoning, code, math | 128K | 文本 | Complex reasoning, code generation, math |
Qwen | Full lineup, multimodal | 最长 1M | 文本 + 多模态 | Multimodal apps, self-host replacement |
Doubao | Chinese writing, dialogue | 256K | 文本 | Chinese content, support chat, marketing copy |
Kimi (Moonshot) | Very long context | 200K+ | 文本 | Long documents, deep research |
Zhipu GLM | Tool calling, agents | 128K–256K | 文本 | Agents and function calling |
Hunyuan | Chinese, WeChat ecosystem | 128K–256K | 文本 + 多模态 | Chinese workloads tied to WeChat |
MiniMax | Text + speech + video | 200K+ | 文本 + 语音 + 视频 | Content creation, voice, video |
Context windows are the published typical ranges per vendor; exact figures depend on the model version you pick.