DeepSeek, Alibaba Qwen, Moonshot Kimi, Zhipu GLM, Baidu ERNIE, ByteDance Doubao, MiniMax, StepFun and Tencent Hunyuan are permanent first-class entries in the model, benchmark, pricing and source-monitoring systems.
ChinaDeepSeekV4 Flash
Best fit: Large-context efficiency and open-model ecosystem
Modalities: Text, reasoning and coding
Context: 1M tokens
Published input: Verify
Published output: Verify
Caveat: Production integrations should use the explicit V4 model names. deepseek-chat and deepseek-reasoner are scheduled to stop working after the retirement deadline.
Open official source↗ (opens in a new tab)ChinaAlibaba CloudQwen3.7-Max
Best fit: Complex multi-step reasoning and coding in the Qwen ecosystem
Modalities: Text, reasoning, coding, tools and structured output
Context: See current Model Studio catalog
Published input: Verify
Published output: Verify
Caveat: Pricing, endpoint names and availability differ between China and international Model Studio regions.
Open official source↗ (opens in a new tab)ChinaMoonshot AIKimi K3
Best fit: Long-context agentic work and software engineering
Modalities: Text, software engineering, knowledge work, deep reasoning and tool calling
Context: 1M tokens
Published input: ¥20/1M
Published output: ¥100/1M
Caveat: Prices are official CNY rates per million tokens and must not be displayed as USD.
Open official source↗ (opens in a new tab)ChinaZhipu AIGLM-5.2
Best fit: Long-horizon coding and autonomous agent workflows
Modalities: Text, reasoning, coding, long-running agents and tools
Context: 1M tokens
Published input: Verify
Published output: Verify
Caveat: Provider capability statements require independent task-specific reproduction.
Open official source↗ (opens in a new tab)ChinaBaiduERNIE 5.0
Best fit: Unified multimodal tasks and Chinese-language applications
Modalities: Unified text, image, video and audio understanding/generation
Context: 128K tokens
Published input: $1.4/1M
Published output: $5.6/1M
Caveat: The displayed API price is from Baidu’s international Qianfan catalog; China-region billing differs.
Open official source↗ (opens in a new tab)ChinaByteDanceDoubao Seed 2.1
Best fit: Production-oriented agent, coding and multimodal tasks
Modalities: General agents, coding and multimodal workflows
Context: See current Volcengine Ark catalog
Published input: Verify
Published output: Verify
Caveat: Volcengine pricing uses tiered regional tables; verify the exact model ID and input-length tier.
Open official source↗ (opens in a new tab)ChinaMiniMaxMiniMax-M3
Best fit: Agentic reasoning, coding and long-context work
Modalities: Text, multimodal chat input, coding, tool use and long-context agents
Context: 1M tokens
Published input: Verify
Published output: Verify
Caveat: Use the current regional pricing table and distinguish API pay-as-you-go from Token Plan subscriptions.
Open official source↗ (opens in a new tab)ChinaStepFunStep 3.7 Flash
Best fit: Low-cost multimodal reasoning and computer-use style tasks
Modalities: Multimodal reasoning, visual interaction, coding and agent workflows
Context: See current StepFun model catalog
Published input: ¥1.35/1M
Published output: ¥8.1/1M
Caveat: Official prices are CNY per million tokens. Confirm current quotas and model lifecycle notices.
Open official source↗ (opens in a new tab)ChinaTencentHunyuan A13B
Best fit: Efficient mixture-of-experts reasoning in the Tencent ecosystem
Modalities: Text, hybrid reasoning, math, science, long documents and agents
Context: 224K max input; 32K max output
Published input: Verify
Published output: Verify
Caveat: Tencent is migrating newer model access toward TokenHub; verify the active platform before integration.
Open official source↗ (opens in a new tab)ChinaAlibaba CloudQwen-Image-3.0
Best fit: Dense layouts, small text, multilingual rendering, interfaces and infographics
Modalities: Text-to-image generation and image editing
Context: Up to 4.5K-token image instructions
Published input: Verify
Published output: Verify
Caveat: Capabilities and examples are provider-published; independent benchmark reproduction is pending.
Open official source↗ (opens in a new tab)ChinaAlibaba CloudQwen-Audio-3.0-TTS Flash / Plus
Best fit: Real-time Flash variant and higher-fidelity Plus variant
Modalities: Text-to-speech, voice cloning, style control and non-verbal tags
Context: 16-language text-to-speech release
Published input: Verify
Published output: Verify
Caveat: The 300ms-level first-packet latency and multilingual results are provider-reported; verify independently for production workloads.
Open official source↗ (opens in a new tab)