AIUpdateWatch Daily Intelligence

AI model capability tracker

A practical global comparison. Main Chinese models are part of the core tracker, with facts, provider claims and AIUpdateWatch interpretation kept separate.

Morning edition: Friday, July 24, 2026 · Data cutoff Jul 23, 2026, 8:30 PM (America/New_York)

Complete current model registry

Prices remain in their source currency. A CNY price is never relabeled as USD, and region-specific access is kept visible.

OriginProviderModelReleasedAvailabilityContextInputOutput
InternationalOpenAIGPT-5.6 Sol2026-07-09API and eligible ChatGPT plansSee current model documentation (opens in a new tab)$5/1M$30/1M
InternationalOpenAIGPT-5.6 Terra2026-07-09API and eligible ChatGPT plansSee current model documentation (opens in a new tab)$2.5/1M$15/1M
InternationalOpenAIGPT-5.6 Luna2026-07-09API and eligible ChatGPT plansSee current model documentation (opens in a new tab)$1/1M$6/1M
InternationalAnthropicClaude Fable 52026-06Claude and API where availableSee current model documentation (opens in a new tab)$10/1M$50/1M
InternationalxAIGrok 4.52026-07-16xAI API and Grok products500K tokens (opens in a new tab)$2/1M$6/1M
InternationalGoogleGemini 3.5 Flash2026-05-19Gemini ecosystem and developer servicesSee current model documentation (opens in a new tab)VerifyVerify
ChinaDeepSeekV4 Flash2026-04-24 previewDeepSeek API via deepseek-v4-flash or deepseek-v4-pro; legacy aliases retire July 24 at 15:59 UTC1M tokens (opens in a new tab)VerifyVerify
ChinaAlibaba CloudQwen3.7-Max2026Alibaba Cloud Model Studio; regional endpoints varySee current Model Studio catalog (opens in a new tab)VerifyVerify
ChinaMoonshot AIKimi K32026-07Kimi API platform; region and account eligibility vary1M tokens (opens in a new tab)¥20/1M¥100/1M
ChinaZhipu AIGLM-5.22026Zhipu BigModel platform; regional access varies1M tokens (opens in a new tab)VerifyVerify
ChinaBaiduERNIE 5.02026-01Baidu Qianfan and ERNIE services; international catalog availability varies128K tokens (opens in a new tab)$1.4/1M$5.6/1M
ChinaByteDanceDoubao Seed 2.12026Volcengine Ark; regional availability variesSee current Volcengine Ark catalog (opens in a new tab)VerifyVerify
ChinaMiniMaxMiniMax-M32026-06-01MiniMax API platform and compatible endpoints1M tokens (opens in a new tab)VerifyVerify
ChinaStepFunStep 3.7 Flash2026StepFun open platform; regional access variesSee current StepFun model catalog (opens in a new tab)¥1.35/1M¥8.1/1M
ChinaTencentHunyuan A13B2025-06-25Tencent Cloud Hunyuan / migration path to TokenHub224K max input; 32K max output (opens in a new tab)VerifyVerify
ChinaAlibaba CloudQwen-Image-3.02026-07-22Qwen and Alibaba Cloud ecosystem; verify regional endpoint availabilityUp to 4.5K-token image instructions (opens in a new tab)VerifyVerify
ChinaAlibaba CloudQwen-Audio-3.0-TTS Flash / Plus2026-07-21Alibaba Cloud Model Studio; rollout and regional availability may vary16-language text-to-speech release (opens in a new tab)VerifyVerify

International providers

OpenAI, Anthropic, Google and xAI

OpenAI

GPT-5.6 Sol

Best fit: Premium reasoning and agentic work

Modalities: Text, vision and tool-enabled workflows

Context: See current model documentation

Important caveat: Launch comparisons are provider-published; confirm task-specific performance independently.

Open official source (opens in a new tab)
OpenAI

GPT-5.6 Terra

Best fit: Mid-tier capability and cost balance

Modalities: Text, vision and tool-enabled workflows

Context: See current model documentation

Important caveat: Use current API documentation for limits and regional availability.

Open official source (opens in a new tab)
OpenAI

GPT-5.6 Luna

Best fit: Lowest-cost member of the GPT-5.6 family

Modalities: Text and general assistant workflows

Context: See current model documentation

Important caveat: Lower price does not imply best performance for every workload.

Open official source (opens in a new tab)
Anthropic

Claude Fable 5

Best fit: Premium long-running agent work

Modalities: Text, vision, coding and long-running agents

Context: See current model documentation

Important caveat: Anthropic documents additional safeguards and retention behavior for Mythos-class traffic.

Open official source (opens in a new tab)
xAI

Grok 4.5

Best fit: Coding, agentic tasks and cost efficiency

Modalities: Text, code, web/X search and tools

Context: 500K tokens

Important caveat: Requests exceeding 200K context use higher prices: $4 input and $12 output per 1M tokens.

Open official source (opens in a new tab)
Google

Gemini 3.5 Flash

Best fit: Fast agentic and multimodal workflows

Modalities: Multimodal, coding and interactive UI generation

Context: See current model documentation

Important caveat: This edition did not capture a stable region-neutral official API price; verify before purchase.

Open official source (opens in a new tab)

China model watch

Main Chinese models tracked every day

DeepSeek, Alibaba Qwen, Moonshot Kimi, Zhipu GLM, Baidu ERNIE, ByteDance Doubao, MiniMax, StepFun and Tencent Hunyuan are permanent first-class entries in the model, benchmark, pricing and source-monitoring systems.

ChinaDeepSeek

V4 Flash

Best fit: Large-context efficiency and open-model ecosystem

Modalities: Text, reasoning and coding

Context: 1M tokens

Published input: Verify

Published output: Verify

Caveat: Production integrations should use the explicit V4 model names. deepseek-chat and deepseek-reasoner are scheduled to stop working after the retirement deadline.

Open official source (opens in a new tab)
ChinaAlibaba Cloud

Qwen3.7-Max

Best fit: Complex multi-step reasoning and coding in the Qwen ecosystem

Modalities: Text, reasoning, coding, tools and structured output

Context: See current Model Studio catalog

Published input: Verify

Published output: Verify

Caveat: Pricing, endpoint names and availability differ between China and international Model Studio regions.

Open official source (opens in a new tab)
ChinaMoonshot AI

Kimi K3

Best fit: Long-context agentic work and software engineering

Modalities: Text, software engineering, knowledge work, deep reasoning and tool calling

Context: 1M tokens

Published input: ¥20/1M

Published output: ¥100/1M

Caveat: Prices are official CNY rates per million tokens and must not be displayed as USD.

Open official source (opens in a new tab)
ChinaZhipu AI

GLM-5.2

Best fit: Long-horizon coding and autonomous agent workflows

Modalities: Text, reasoning, coding, long-running agents and tools

Context: 1M tokens

Published input: Verify

Published output: Verify

Caveat: Provider capability statements require independent task-specific reproduction.

Open official source (opens in a new tab)
ChinaBaidu

ERNIE 5.0

Best fit: Unified multimodal tasks and Chinese-language applications

Modalities: Unified text, image, video and audio understanding/generation

Context: 128K tokens

Published input: $1.4/1M

Published output: $5.6/1M

Caveat: The displayed API price is from Baidu’s international Qianfan catalog; China-region billing differs.

Open official source (opens in a new tab)
ChinaByteDance

Doubao Seed 2.1

Best fit: Production-oriented agent, coding and multimodal tasks

Modalities: General agents, coding and multimodal workflows

Context: See current Volcengine Ark catalog

Published input: Verify

Published output: Verify

Caveat: Volcengine pricing uses tiered regional tables; verify the exact model ID and input-length tier.

Open official source (opens in a new tab)
ChinaMiniMax

MiniMax-M3

Best fit: Agentic reasoning, coding and long-context work

Modalities: Text, multimodal chat input, coding, tool use and long-context agents

Context: 1M tokens

Published input: Verify

Published output: Verify

Caveat: Use the current regional pricing table and distinguish API pay-as-you-go from Token Plan subscriptions.

Open official source (opens in a new tab)
ChinaStepFun

Step 3.7 Flash

Best fit: Low-cost multimodal reasoning and computer-use style tasks

Modalities: Multimodal reasoning, visual interaction, coding and agent workflows

Context: See current StepFun model catalog

Published input: ¥1.35/1M

Published output: ¥8.1/1M

Caveat: Official prices are CNY per million tokens. Confirm current quotas and model lifecycle notices.

Open official source (opens in a new tab)
ChinaTencent

Hunyuan A13B

Best fit: Efficient mixture-of-experts reasoning in the Tencent ecosystem

Modalities: Text, hybrid reasoning, math, science, long documents and agents

Context: 224K max input; 32K max output

Published input: Verify

Published output: Verify

Caveat: Tencent is migrating newer model access toward TokenHub; verify the active platform before integration.

Open official source (opens in a new tab)
ChinaAlibaba Cloud

Qwen-Image-3.0

Best fit: Dense layouts, small text, multilingual rendering, interfaces and infographics

Modalities: Text-to-image generation and image editing

Context: Up to 4.5K-token image instructions

Published input: Verify

Published output: Verify

Caveat: Capabilities and examples are provider-published; independent benchmark reproduction is pending.

Open official source (opens in a new tab)
ChinaAlibaba Cloud

Qwen-Audio-3.0-TTS Flash / Plus

Best fit: Real-time Flash variant and higher-fidelity Plus variant

Modalities: Text-to-speech, voice cloning, style control and non-verbal tags

Context: 16-language text-to-speech release

Published input: Verify

Published output: Verify

Caveat: The 300ms-level first-packet latency and multilingual results are provider-reported; verify independently for production workloads.

Open official source (opens in a new tab)

How to choose

Quality first

Use a small test set representing your real tasks. Do not choose from a single leaderboard score.

Cost first

Include currency, output length, caching, tools, retries, batch discounts and long-context surcharges.

Privacy first

Review retention, regional processing, enterprise controls, export restrictions and whether a local model is practical.