Chinese AI Model Rankings
As of 2026-09-22, Claude Opus 5.5 leads the Artificial Analysis Intelligence Index at 58; the highest-placed Chinese model is MiMo-V2.6-Pro at 46 (#10). 82 of the 100 models listed can be called through ChinaAPI.
Where the Chinese AI models ChinaAPI serves stand next to the international frontier, quoted from public leaderboards. Snapshot of 2026-09-22; roster last maintained 2026-09-22 (moves when a model or source is added). Every figure is reproduced from the board named in its column and links back to it. ChinaAPI runs no evaluation of its own and computes no composite score from these numbers. See how this list is built, the price of every model marked available, and the same data as JSON at /api/model-rankings.
Sources
- LMArena Text Arena — LMArena (arena.ai), Arena score (Elo), live board, retrieved 2026-09-22
- LiveBench — LiveBench (sponsored by Abacus.AI), Overall, LiveBench-2026-06-25, retrieved 2026-09-22
- Artificial Analysis Intelligence Index — Artificial Analysis, Intelligence Index, Intelligence Index v4.3.2, read 2026-09-22; its figures are not comparable with the v4.3 figures this page quoted from its 2026-09-12 read. v4.3.1 refreshed the pairwise judge panels of AA-Briefcase and GDPval-AA, and v4.3.2 moved GDPval-AA to v2.1 (Elo scale anchored to DeepSeek V4.1 Flash (max) at 1600, fitted with a Crowd-BT model) and refitted AA-Briefcase v1.1 with Crowd-BT, so rows moved without any model changing. v4.3 had already swapped τ³-Banking out of the index for AutomationBench-AA and Terminal-Bench v2.1 for v4.0. The leaderboard's default Status: Current filter hides rows Artificial Analysis marks deprecated; they still render with Status set to All and are re-read from that view like any other row (CA-948). A trailing * is reproduced as the board prints it: the board's data marks that score as estimated (intelligenceIndexIsEstimated), i.e. not every evaluation in the index had been run for that model, retrieved 2026-09-22
- SuperCLUE 智能指数 — SuperCLUE, 总分 (overall), 2026-07 evaluation, published 2026-08-06. The homepage reads this edition from its data file (/data/generalboard/2026年7月.xlsx, sheet 总排行榜, last modified 2026-09-11), which also carries rows evaluated after publication, each with its own evaluation publication date (the latest 2026-09-11); the page's Latest Update line refers to a different sub-board, retrieved 2026-09-22
- Epoch Capabilities Index (ECI) — Epoch AI, General ECI, live board, retrieved 2026-09-22, CC BY 4.0, as labelled on the page (https://creativecommons.org/licenses/by/4.0/)
- Vals Index — Vals AI, Accuracy (GDP-weighted finance, coding and legal), V2 (released 2026-08-13), page states updated 9/21/2026, retrieved 2026-09-22
- Vendor's own evaluation card — the model's own vendor, varies; the benchmark is named in each row, self-reported, retrieved 2026-09-22
- LMArena Text-to-Image Arena — LMArena (arena.ai), Arena score (Elo), live board, last updated 2026-09-07, retrieved 2026-09-12
- LMArena Image Edit Arena — LMArena (arena.ai), Arena score (Elo), live board, last updated 2026-09-07, retrieved 2026-09-12
- LMArena Text-to-Video Arena — LMArena (arena.ai), Arena score (Elo), live board, last updated 2026-09-04, retrieved 2026-09-12
- LMArena Image-to-Video Arena — LMArena (arena.ai), Arena score (Elo), live board, last updated 2026-09-02, retrieved 2026-09-12
- LMArena Text Arena · Chinese — LMArena (arena.ai), Arena score (Elo), live board, Chinese-language category (prompts in Chinese), retrieved 2026-09-22
- LMArena Text Arena · Coding — LMArena (arena.ai), Arena score (Elo), live board, Coding category, retrieved 2026-09-22
- LiveBench · Coding — LiveBench (sponsored by Abacus.AI), Coding, LiveBench-2026-06-25, retrieved 2026-09-22
- LiveBench · Agentic Coding — LiveBench (sponsored by Abacus.AI), Agentic Coding, LiveBench-2026-06-25, retrieved 2026-09-22
- Artificial Analysis · Output Speed — Artificial Analysis, Median output tokens/s, live board; measured by Artificial Analysis against the vendor's own API, not through ChinaAPI. Rows Artificial Analysis marks deprecated are read with the leaderboard's Status filter set to All (CA-948), retrieved 2026-09-22
- Artificial Analysis · GDPval-AA v2.1 — Artificial Analysis, Agentic Real-World Work Tasks, (Elo-500)/2000 (%), GDPval-AA v2.1, live board, read 2026-09-22. 220 agentic task-completion tasks with file outputs, ranked pairwise by a panel of three frontier LLM judges; since v2.1 the Elo scale is anchored to DeepSeek V4.1 Flash (max) at 1600 and fitted with a Crowd-BT model, so its figures are not comparable with the v2 figures this page quoted from its 2026-09-12 read. The column prints clamp((Elo-500)/2000). It carries 10% of the Agents category inside Intelligence Index v4.3.2, so this column is a component of the board that orders the table, not an independent operator. Rows Artificial Analysis marks deprecated are read with the leaderboard's Status filter set to All (CA-948). The column is hidden until the leaderboard's Intelligence group is expanded, retrieved 2026-09-22
- Artificial Analysis · Terminal-Bench v4.0 — Artificial Analysis, Agentic Coding & Terminal Use (%), Terminal-Bench 4.0 (66 tasks), run by Artificial Analysis with the mini-swe-agent harness, pass@1 averaged over 3 repeats per task; read 2026-09-22. When this column replaced the benchmark's own board on 2026-09-07 the methodology page named the harness as mini-SWE-agent v2.4.6; on 2026-09-21 it and the evaluation page name mini-swe-agent without a version, and every figure quoted before then read back unchanged. One harness for every model, so the column compares models rather than model-plus-best-scaffold — which is why it replaced the benchmark's own board, whose rows are agent × model and covered one Chinese model out of 34 (CA-788, overturning CA-734). It carries 10% of the Coding category inside Intelligence Index v4.3.2, so this column is a component of the board that orders the table, not an independent operator. Rows Artificial Analysis marks deprecated are read with the leaderboard's Status filter set to All (CA-948). The column is hidden until the leaderboard's Intelligence group is expanded, retrieved 2026-09-22
- Artificial Analysis · MLCR-AA — Artificial Analysis, Medical Long Context Reasoning (%), MLCR-AA, live board, read 2026-09-22. Artificial Analysis marks it a standalone evaluation that is NOT part of Intelligence Index v4.3 — unlike the GDPval-AA v2 and Terminal-Bench v4.0 columns, this one is not a component of the board that orders the table, only the same operator. The underlying benchmark is Wisedocs' open MLCR (Wisedocs-AI/medical-long-context-reasoning): synthetic medical records of roughly 25,000-64,000 tokens, graded across six tiers from locating a single fact to expert-level clinical synthesis. It measures multi-document reasoning over long records, not medical capability in general. The main leaderboard carries no column for it; the figures were first read off the evaluation page's chart (2026-09-07) and since 2026-09-22 from the per-model pages' payload, which carries every model the chart can show (CA-956), retrieved 2026-09-22
- Artificial Analysis Text To Image — Artificial Analysis, Elo, live board; exact row settings retained, retrieved 2026-09-13
- Artificial Analysis Editing — Artificial Analysis, Elo, live board; exact row settings retained, retrieved 2026-09-13
- Artificial Analysis Text To Video · With Audio — Artificial Analysis, Elo, live board; with audio pool; exact row settings retained, retrieved 2026-09-13
- Artificial Analysis Image To Video · With Audio — Artificial Analysis, Elo, live board; with audio pool; exact row settings retained, retrieved 2026-09-13
Text and multimodal models
Ordered by Artificial Analysis Intelligence Index; a model with no figure on that board follows the ones that have one, because absent from a board is not the same as last on it.
| # | Model | Vendor | Origin | Released | On ChinaAPI | LMArena Text Arena | LiveBench | Artificial Analysis Intelligence Index | SuperCLUE 智能指数 | Epoch Capabilities Index (ECI) | Vals Index | Vendor's own evaluation card | LMArena Text Arena · Chinese | LMArena Text Arena · Coding | LiveBench · Coding | LiveBench · Agentic Coding | Artificial Analysis · Output Speed | Artificial Analysis · GDPval-AA v2.1 | Artificial Analysis · Terminal-Bench v4.0 | Artificial Analysis · MLCR-AA | Notes |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 1 | Claude Opus 5.5 | Anthropic | International | 2026-09-22 | yes | — | 83.2 | 58 | — | — | 66.16 ±1.00 | — | — | — | 89.3 | 71.7 | — | 67 | 60 | — | Released 2026-09-22 per Anthropic's announcement (Introducing Claude Opus 5.5), the first model in the Claude 5.5 family. Artificial Analysis printed no output speed for its Claude Opus 5.5 (max with fallback) row on 2026-09-22 (the column reads --), so no speed figure is quoted; the speeds of its lower-effort rows are not substituted. |
| 2 | Claude Fable 5.1 | Anthropic | International | 2026-09-01 | yes | 1498 ±8 | 83.4 | 53 | — | 165 (162 - 170) | 68.83 ±1.08 | — | 1592 ±32 | 1519 ±18 | 86.4 | 66.1 | 65 | 62 | 52 | 71.1 | Released 2026-09-01 per Anthropic's announcement (Introducing Claude Fable 5.1 and Claude Mythos 5.1); on sale here since 2026-09-02. LMArena's row is the max-effort build, 5,783 votes as of 2026-09-22 and not marked preliminary; the same row on the Chinese category stands on 369 votes. Epoch's fit as read on 2026-09-22 scores it 165.00 and Claude Fable 5 163.60. |
| 3 | GPT-6 Astra | OpenAI | International | 2026-09-03 | yes | 1480 ±12 | 82.2 | 53 | — | 167 (163 - 172) | 66.61 ±1.09 | — | — | 1543 ±23 | 80.4 | 57.3 | 58 | 52 | 59 | 35.0 | Released 2026-09-03 per OpenAI's announcement (GPT-6 Astra: A new generation of intelligence). LMArena added gpt-6-astra-max to its text board on 2026-09-11 per its leaderboard changelog; the row stands on 2,693 votes as of 2026-09-22, and the Chinese category carries no row for it. On sale here since 2026-09-05. |
| 4 | Claude Opus 5 | Anthropic | International | 2026-07-24 | yes | 1493 ±4 | 80.1 | 51 | — | 163 (160 - 167) | 67.21 ±0.98 | — | 1559 ±12 | 1533 ±7 | 81.4 | 65.2 | 54 | 60 | 49 | 55.6 | |
| 5 | Claude Fable 5 | Anthropic | International | 2026-06-09 | yes | 1506 ±5 | 83.0 | 50 | — | 164 (161 - 168) | 66.04 ±1.03 | — | 1553 ±14 | 1552 ±7 | 86.0 | 62.2 | 61 | 55 | 42 | 64.4 | Artificial Analysis marked its Claude Fable 5 (with fallback) row deprecated between the 2026-09-12 and 2026-09-21 reads; the leaderboard hides such rows under its default Status: Current filter, so this row's leaderboard figures are read with the Status filter set to All. LMArena relabelled this row from claude-fable-5 to claude-fable-5-high between 2026-09-06 and 2026-09-18: the Internet Archive's copies of the board show the same model key (claude-fable-5-text) with its vote count running on from 27,189 to 30,057, so the entry follows the new label on all three LMArena text boards. |
| 6 | GPT-6 Sol | OpenAI | International | 2026-09-22 | yes | — | — | 48 | — | — | — | — | — | — | — | — | 104 | 49 | 44 | — | Released 2026-09-22 per OpenAI's announcement (Introducing GPT-6 Sol and Luna) and the API changelog entry of the same day. |
| 7 | Muse Spark 1.3 | Meta | International | 2026-09-02 | no | 1493 ±9 | 81.6 | 48 | — | 157 (155 - 159) | 60.31 ±1.08 | — | 1541 ±34 | 1537 ±16 | 81.1 | 64.1 | 219 | 59 | 33 | 43.3 | Released 2026-09-02 per Meta AI Research's announcement (Introducing Muse Spark 1.3). Epoch, which carried no row for it as of 2026-09-05, lists it as Muse Spark 1.3 (dated 2026-09-02) in the 2026-09-22 read. LMArena added muse-spark-1.3-max to its text board on 2026-09-13 per its leaderboard changelog; the row stands on 4,723 votes as of 2026-09-22. Not on sale here as of 2026-09-22. |
| 8 | GPT-5.6 Sol | OpenAI | International | 2026-07-09 | yes | 1483 ±5 | 81.0 | 47 | — | 162 (160 - 166) | 63.71 ±1.06 | — | 1539 ±15 | 1528 ±8 | 83.9 | 56.2 | 73 | 54 | 40 | 26.1 | OpenAI announced a limited partner preview of the GPT-5.6 series (Sol, Terra and Luna) on 2026-06-26 and made the models generally available on 2026-07-09, the post the preview page links as the launch; release_date follows the launch. |
| 9 | Grok 4.7 | xAI | International | 2026-09-21 | no | — | — | 46 | — | — | 60.20 ±1.08 | — | — | — | — | — | 39 | 60 | 26 | 15.0 | Released 2026-09-21 per the vendor's announcement, which now publishes under the SpaceXAI name; the row keeps the xAI label of the earlier Grok rows. Artificial Analysis lists Grok 4.7 (xhigh) and Grok 4.7 (high), both printing 46; the xhigh row is quoted because it is the highest effort setting and the higher score before rounding (46.4 against 46.3). |
| 10 | MiMo-V2.6-Pro | Xiaomi | China | 2026-09-22 | yes | — | — | 46 | — | — | — | — | — | — | — | — | 55 | 59 | 35 | — | Released and open-sourced 2026-09-22 per Xiaomi's release record (MiMo-V2.6 Series Release). |
| 11 | GLM-5.3 | Zhipu AI (Z.ai) | China | 2026-08-14 | yes | 1483 ±6 | 76.1 | 45 | 71.29 | 156 (154 - 158) | 56.97 ±1.35 | — | 1525 ±22 | 1524 ±12 | 79.0 | 60.9 | 53 | 57 | 42 | 48.3 | Released 2026-08-14; served here via Zhipu's official channels since 2026-08-18. Independent boards picked it up between the 2026-08-19 and 2026-08-26 snapshots, so the vendor's own DeepSWE figure it carried before has been replaced by their readings; the LMArena row stands on 10,960 votes as of 2026-09-22. Epoch, which carried no row for it as of 2026-09-05, lists it as GLM-5.3 (dated 2026-08-14) in the 2026-09-22 read, and SuperCLUE lists GLM-5.3(max) with an evaluation published 2026-09-10. Vals lists it as GLM 5.3, released 2026-08-18 per its model page, the date of Z.ai's release notes. |
| 12 | Qwen3.8-Max-0902 | Alibaba | China | 2026-09-02 | yes | — | — | 45 | 72.62 | 155 (153 - 157) | — | — | — | — | — | — | 39 | 58 | 39 | 20.0 | Alibaba's 2026-09-02 snapshot of Qwen3.8-Max (alias qwen3.8-max-2026-09-02), listed on Model Studio on 2026-09-02 as an upgraded snapshot of qwen3.8-max; Epoch dates it 2026-09-01. Alibaba Cloud's update notices of 2026-09-02 (www.aliyun.com/notice/118616 and the international Model Studio notice Update Notice for Qwen3.8-Max Models) state that from 10:00 Beijing time on 2026-09-05, subject to the actual change time, the qwen3.8-max model name automatically moved to this snapshot with billing unchanged. Alibaba published no completion notice; Artificial Analysis has since marked the earlier build deprecated. This row therefore links to qwen3.8-max, the name on sale here; the same snapshot is also sold under its pinned name qwen3.8-max-0902. Artificial Analysis and Epoch list it as Qwen3.8 Max (0902), SuperCLUE as Qwen3.8-Max-0902(max) (evaluation published 2026-09-10). LMArena added qwen3.8-max-0902 only to its Code Arena (changelog, 2026-09-01), which this page does not cite, and neither LiveBench nor Vals carries a 0902 row; the 2026-08-02 build's figures on the qwen3.8-max row are not reused here. |
| 13 | Grok 4.6 | xAI | International | 2026-08-12 | no | 1456 ±6 | 78.0 | 44 | — | 156 (155 - 159) | 59.17 ±1.21 | — | 1517 ±18 | 1506 ±10 | 76.8 | 57.0 | 58 | 55 | 21 | 12.2 | Released 2026-08-12. LMArena labels this vendor SpaceXAI. Its row carried the Preliminary badge on 3,453 votes as of 2026-09-05 and no longer does (15,521 votes as of 2026-09-22). |
| 14 | Kimi K3 | Moonshot AI | China | 2026-07-16 | yes | 1485 ±5 | 79.2 | 44 | 70.68 | 158 (155 - 161) | 57.81 ±1.06 | — | 1535 ±16 | 1538 ±9 | 81.4 | 62.2 | 37 | 51 | 13 | 38.3 | |
| 15 | Step 5 Preview | StepFun | China | 2026-09-20 | yes | — | — | 44 | — | — | — | — | — | — | — | — | 83 | 53 | 33 | 16.7 | StepFun's announcement page prints no date; it says the model is released today and cites Artificial Analysis results published as of 2026-09-20, and reports dated that day already quote the release, which fixes the release date. The vendor states open weights follow on 2026-10-15. |
| 16 | GLM-5.3-Flash | Zhipu AI (Z.ai) | China | 2026-08-26 | yes | 1475 ±7 | 71.6 | 42 | 68.10 | 152 (150 - 154) | 47.22 ±1.45 | — | 1529 ±25 | 1525 ±12 | 79.0 | 56.8 | 65 | 57 | 33 | 51.1 | Released 2026-08-26 per Z.ai's own release notes; the first natively multimodal model in the GLM-5 line and its low-cost tier, served here on Zhipu's official channels. LMArena, LiveBench and Artificial Analysis all picked it up within days of release; the LMArena row is a ranked one standing on 10,038 votes as of 2026-09-22, carrying neither the board's Preliminary badge nor an AutoEval marker. Epoch, SuperCLUE and Vals carried no row for it as of 2026-08-31. In the 2026-09-22 read Epoch lists it as GLM-5.3-Flash dated 2026-08-20, six days before Z.ai's release notes; the benchmark runs behind that figure are on the glm-5.3-flash model itself (Epoch's model versions glm-5.3-flash_max and glm-5.3-flash_high), so only the date differs. SuperCLUE lists GLM-5.3-Flash(max) with an evaluation published 2026-09-10, and Vals lists it as GLM 5.3 Flash, released 2026-08-26 per its model page. GLM-5.3's figures are not reused here. |
| 17 | GPT-5.6 Terra | OpenAI | International | 2026-07-09 | yes | 1466 ±5 | 77.9 | 42 | — | 159 (157 - 162) | 59.59 ±1.33 | — | 1516 ±14 | 1518 ±8 | 78.2 | 54.9 | 82 | 47 | 35 | 31.7 | OpenAI announced a limited partner preview of the GPT-5.6 series (Sol, Terra and Luna) on 2026-06-26 and made the models generally available on 2026-07-09, the post the preview page links as the launch; release_date follows the launch. |
| 18 | Gemini 3.8 Flash | International | 2026-09-02 | yes | 1493 ±9 | 75.8 | 41 | — | 157 (155 - 161) | 62.25 ±1.01 | — | 1543 ±32 | 1535 ±16 | 72.5 | 54.2 | 297 | 46 | 20 | 21.7 | Released 2026-09-02 per Google's announcement (Introducing Gemini 3.8 Flash and 3.8 Flash Cyber). LMArena marks the row preliminary (5,076 votes as of 2026-09-22). Epoch, which carried no row for it as of 2026-09-05, lists it as Gemini 3.8 Flash (dated 2026-09-02) in the 2026-09-22 read. Confirmed on sale in the public catalogue read on 2026-09-12. | |
| 19 | Muse Spark 1.2 | Meta | International | 2026-08-05 | no | 1500 ±11 | 78.0 | 40 | — | 155 (153 - 158) | 57.05 ±1.12 | — | 1528 ±38 | 1536 ±20 | 77.5 | 57.6 | 172 | 49 | 7 | 31.1 | Epoch added a Muse Spark 1.2 row (dated 2026-08-05) between the 2026-08-31 and 2026-09-05 snapshots. Artificial Analysis marked its Muse Spark 1.2 (xhigh) row deprecated between the 2026-09-12 and 2026-09-21 reads; the leaderboard hides such rows under its default Status: Current filter, so this row's leaderboard figures are read with the Status filter set to All. |
| 20 | Qwen3.8-2.4T-A95B | Alibaba | China | — | yes | — | — | 40 | — | — | — | — | — | — | — | — | 38 | 55 | 11 | 0.0 | The open-weight Qwen3.8-2.4T-A95B release has its own Artificial Analysis row. Qwen identifies Qwen3.8-Max as a production version based on this model with additional features; the two catalogue models remain separate here, and Max scores from other boards are not reused. Artificial Analysis prints MLCR-AA 0.0 for this release, with no breakdown, against 19.4 and 20.0 for the two Qwen3.8-Max builds. |
| 21 | Qwen3.8-Flash | Alibaba | China | 2026-08-26 | yes | — | 76.2 | 40 | 66.41 | — | — | — | — | — | 72.6 | 61.6 | 54 | 56 | 25 | — | The boards list this model under its open-weight release name, Qwen3.8-Flash-Next (released 2026-08-26 per the Qwen team). The vendor's announcement states that the production version, with a 1M default context and built-in tools, is served under the name Qwen3.8-Flash, which is the qwen3.8-flash API on sale here since 2026-09-04; the two rows are joined on that statement, not on the row name. Artificial Analysis now prints 40 without an estimated marker (read 2026-09-21). SuperCLUE lists it under the API name itself, Qwen3.8-Flash(max), evaluated through the API with an evaluation published 2026-09-10, so that figure needs no join. Neither Epoch, LMArena nor Vals carried a row as of 2026-09-05. |
| 22 | Qwen3.8-Max | Alibaba | China | 2026-08-02 | no | 1481 ±6 | 78.5 | 40 | 71.48 | 157 (155 - 159) | 51.84 ±1.29 | — | 1538 ±18 | 1522 ±9 | 72.9 | 64.6 | 38 | 55 | 19 | 19.4 | The 2026-08-02 build of Qwen3.8-Max (Artificial Analysis dates it 2026-08-03), which replaced Qwen3.8-Max-Preview (retired 2026-08-05). Its release_date follows Epoch, which re-dated this build from 2026-07-19, the Preview's announcement at WAIC, to 2026-08-02 between its 2026-08-25 and 2026-09-03 releases; Model Studio listed qwen3.8-max on 2026-08-02 as well. Alibaba Cloud's update notices of 2026-09-02 (www.aliyun.com/notice/118616 and the international Model Studio notice Update Notice for Qwen3.8-Max Models) state that from 10:00 Beijing time on 2026-09-05, subject to the actual change time, the qwen3.8-max model name automatically moved to the qwen3.8-max-0902 snapshot with billing unchanged. The name on sale here therefore now serves the Qwen3.8-Max-0902 row, and Model Studio lists no snapshot ID for this build, so it can no longer be called there. Every figure on this row belongs to this build: Artificial Analysis's Qwen3.8 Max (slug qwen3-8-max-0803), Epoch's Qwen 3.8 Max (dated 2026-08-02), SuperCLUE's Qwen3.8-Max(max) (evaluation published 2026-08-06), LMArena's qwen3.8-max (added to its Text Arena on 2026-08-02 per its changelog), LiveBench's Qwen 3.8 Max (quoted at the same figures since 2026-08-10, before the snapshot existed) and Vals's Qwen 3.8 Max (released 2026-08-03 per its model page). Artificial Analysis marked its Qwen3.8 Max row deprecated between the 2026-09-12 and 2026-09-21 reads; the leaderboard hides such rows under its default Status: Current filter, so this row's leaderboard figures are read with the Status filter set to All. |
| 23 | DeepSeek V4.1 Flash | DeepSeek | China | 2026-09-10 | yes | — | 81.1 | 39 | 71.81 | 155 (149 - 158) | 57.86 ±1.16 | — | — | — | 80.0 | 77.3 | 220 | 55 | 27 | 22.8 | Released 2026-09-10 per DeepSeek's official announcement and served by the canonical deepseek-flash API model. The vendor retired V4 Flash and V4 Flash Vision Exp that day and temporarily routes their old identifiers to V4.1 Flash; their historical leaderboard rows remain separate and none of their figures are reused here. Artificial Analysis, LiveBench and Vals all published distinct V4.1 Flash rows by 2026-09-12; Epoch (DeepSeek V4.1 Flash, dated 2026-09-09) and SuperCLUE (DeepSeek-V4.1-Flash(max), evaluation published 2026-09-11) carry one as of the 2026-09-22 read. |
| 24 | Gemini 3.7 Flash | International | 2026-08-13 | yes | 1490 ±8 | 78.8 | 39 | — | 158 (156 - 161) | 59.31 ±1.06 | — | 1556 ±30 | 1521 ±15 | 78.9 | 58.3 | 281 | 44 | 14 | 15.0 | Released 2026-08-13. LMArena marks the row preliminary (5,640 votes as of 2026-09-22). Artificial Analysis marked its Gemini 3.7 Flash (high) row deprecated between the 2026-09-12 and 2026-09-21 reads; the leaderboard hides such rows under its default Status: Current filter, so this row's leaderboard figures are read with the Status filter set to All. | |
| 25 | Grok 4.5 | xAI | International | 2026-07-08 | no | 1468 ±5 | 75.8 | 39 | — | 154 (152 - 156) | 51.53 ±1.32 | — | 1512 ±14 | 1518 ±7 | 68.6 | 56.5 | 55 | 44 | 11 | 15.0 | LMArena labels this vendor SpaceXAI. Artificial Analysis marked its Grok 4.5 (high) row deprecated between the 2026-09-12 and 2026-09-21 reads; the leaderboard hides such rows under its default Status: Current filter, so this row's leaderboard figures are read with the Status filter set to All. Released 2026-07-08 per xAI's announcement as published that day and its API release notes; the live announcement page now shows Jul 16, the date it was edited when the model reached the EU. |
| 26 | GPT-5.6 Luna | OpenAI | International | 2026-07-09 | yes | 1452 ±5 | 73.6 | 37 | — | 156 (154 - 159) | 59.88 ±1.09 | — | 1476 ±14 | 1498 ±8 | 82.9 | 48.4 | 144 | 47 | 12 | 19.4 | The low-cost tier of the GPT-5.6 line. OpenAI announced a limited partner preview of the GPT-5.6 series (Sol, Terra and Luna) on 2026-06-26 and made the models generally available on 2026-07-09, the post the preview page links as the launch; release_date follows the launch. |
| 27 | GPT-6 Luna | OpenAI | International | 2026-09-22 | yes | — | — | 37 | — | — | — | — | — | — | — | — | 157 | 43 | 13 | — | Released 2026-09-22 per OpenAI's announcement (Introducing GPT-6 Sol and Luna) and the API changelog entry of the same day. |
| 28 | Agnes 3.0 Flash | Agnes AI | China | — | yes | — | — | 36* | — | — | — | — | — | — | — | — | — | 53 | 7 | — | Artificial Analysis lists Agnes 3.0 Flash separately from the 2.5 family. Its Intelligence Index is currently estimated, so the trailing * is retained exactly as the board prints it. Artificial Analysis no longer printed an output-speed value for this row on 2026-09-21 (the column reads --), so the 238 tokens/s figure was removed rather than retained without a live source; the index, GDPval-AA and Terminal-Bench figures are from the 2026-09-21 board read. |
| 29 | DeepSeek V4 Pro 0813 | DeepSeek | China | 2026-08-13 | yes | 1463 ±7 | 77.4 | 36 | 70.10 | 155 (154 - 157) | 52.37 ±1.14 | — | 1494 ±25 | 1503 ±12 | 77.2 | 54.9 | 68 | 47 | 14 | 17.8 | DeepSeek's change log dates this build's GA release to 2026-08-13 and states that the deepseek-v4-pro model name now serves it. Epoch added a DeepSeek V4 Pro 0813 row between the 2026-08-26 and 2026-08-31 snapshots, so the undated DeepSeek-V4-Pro figure is no longer borrowed for it. LMArena's 0813 row stands on 9,008 human votes as of 2026-09-22. SuperCLUE announced on 2026-08-26 that it had added the official DeepSeek-V4-Pro; in the July 2026 edition's data file read on 2026-09-22 that row is labelled DeepSeek-V4-Pro-0813(max), with an evaluation published 2026-08-18, and the 64.40 row the board had labelled DeepSeek-V4-Pro(max) 预览版, now plain DeepSeek-V4-Pro(max), is the preview build, quoted on the deepseek-v4-pro row rather than here. |
| 30 | DeepSeek V4 Flash Vision Exp | DeepSeek | China | 2026-08-21 | no | — | 76.8 | 35 | — | — | — | — | — | — | 68.2 | 65.1 | 231 | 52 | 12 | — | DeepSeek's experimental vision tier, announced in the vendor's change log on 2026-08-21 as model='deepseek-v4-flash-vision-exp'. LiveBench names its row with the vendor's -Exp suffix and Artificial Analysis drops it, but DeepSeek publishes only one vision tier, so both rows are this model. LMArena, Vals, SuperCLUE and Epoch carry no row for it, and the sibling flash figures are not reused. DeepSeek retired this tier on 2026-09-10 and its compatibility ID now routes to V4.1 Flash. It is absent from the public catalogue read on 2026-09-12 and is retained here as a historical evaluation row without an on-sale badge. |
| 31 | DeepSeek V4 Flash 0731 | DeepSeek | China | 2026-07-31 | no | — | 74.2 | 34 | 65.60 | 155 (152 - 156) | 53.57 ±1.21 | — | — | — | 75.0 | 46.8 | 221 | 46 | 12 | 13.3 | Historical 2026-07-31 V4 Flash build. DeepSeek retired V4 Flash on 2026-09-10 and its compatibility ID now routes to V4.1 Flash; this historical score row is not the currently served build and is not marked on sale. Artificial Analysis marked its DeepSeek V4 Flash 0731 (max) row deprecated between the 2026-09-12 and 2026-09-21 reads; the leaderboard hides such rows under its default Status: Current filter, so this row's leaderboard figures are read with the Status filter set to All. |
| 32 | GLM-5.2 | Zhipu AI (Z.ai) | China | 2026-06-16 | yes | 1472 ±5 | 73.2 | 34 | 63.27 | 152 (150 - 154) | 53.12 ±1.18 | — | 1517 ±13 | 1510 ±7 | 79.7 | 51.8 | 68 | 43 | 1 | 7.2 | Artificial Analysis marked its GLM-5.2 (max) row deprecated between the 2026-09-12 and 2026-09-21 reads; the leaderboard hides such rows under its default Status: Current filter, so this row's leaderboard figures are read with the Status filter set to All. |
| 33 | Gemini 3.6 Flash | International | 2026-07-21 | yes | 1480 ±5 | 73.6 | 34 | — | 154 (153 - 156) | 55.35 ±1.09 | — | 1536 ±15 | 1517 ±8 | 77.9 | 43.4 | 180 | 38 | 7 | 14.4 | Artificial Analysis marked its Gemini 3.6 Flash row deprecated between the 2026-09-12 and 2026-09-21 reads; the leaderboard hides such rows under its default Status: Current filter, so this row's leaderboard figures are read with the Status filter set to All. | |
| 34 | Qwen3.8-27B | Alibaba | China | 2026-08-14 | yes | 1437 ±6 | 75.3 | 34 | — | 149 (147 - 151) | 48.48 ±1.45 | — | 1496 ±23 | 1499 ±11 | 75.7 | 61.4 | 41 | 45 | 6 | 21.7 | Open-weight Qwen3.8 model, released 2026-08-14 per the Qwen team's release notes (Hugging Face and ModelScope); on sale here since 2026-09-04. Epoch, which carried no row for it as of 2026-09-05, lists it as Qwen 3.8 27B (dated 2026-08-14) in the 2026-09-22 read. Artificial Analysis lists several effort settings; the first-listed xhigh row is quoted rather than a lower-effort sibling. |
| 35 | DeepSeek V4 Pro | DeepSeek | China | 2026-04-24 | no | 1457 ±4 | — | 30 | 64.40 | 149 (147 - 151) | 42.89 ±1.19 | — | 1493 ±12 | 1501 ±6 | — | — | 86 | 32 | 15 | 21.1 | The pre-0813 build; the deepseek-v4-pro API moved to DeepSeek-V4-Pro-0813 on 2026-08-12. LiveBench dropped this row in favour of the dated builds between the 2026-08-19 and 2026-08-26 snapshots, so its 71.6 is no longer quoted. SuperCLUE's DeepSeek-V4-Pro(max) row (64.40), which its July 2026 edition labelled 预览版, is quoted here from CA-971 on, overturning CA-605. CA-605 left it out because LMArena lists deepseek-v4-pro and deepseek-v4-pro-high-preview separately, but LMArena's own data gives those two rows the same checkpoint key (deepseek-v4-pro-ch1-text and deepseek-v4-pro-ch1-thinking-text) and links both to DeepSeek's 2026-04-24 announcement: they are one build in two modes, which DeepSeek's change log calls V4-Pro-Preview. Before the 2026-08-13 GA that build was the only one the deepseek-v4-pro API served, so an API evaluation published in the July edition can only be this row. SuperCLUE's data file has since dropped the 预览版 label and dates the row 2026.9.10, keeping the July figure. Artificial Analysis marked its DeepSeek V4 Pro (max) row deprecated between the 2026-09-12 and 2026-09-21 reads; the leaderboard hides such rows under its default Status: Current filter, so this row's leaderboard figures are read with the Status filter set to All. |
| 36 | Gemini 3.1 Pro | International | 2026-02-19 | no | 1487 ±3 | 77.0 | 30 | — | 155 (153 - 158) | 41.90 ±1.17 | — | 1530 ±8 | 1520 ±5 | 76.5 | 44.1 | 115 | 14 | 4 | 15.6 | ||
| 37 | MiniMax-M3 | MiniMax | China | 2026-06-01 | yes | 1441 ±4 | 67.3 | 29 | 56.90 | 147 (143 - 150) | 42.72 ±1.20 | — | 1471 ±12 | 1496 ±6 | 68.2 | 40.7 | 106 | 37 | 2 | 17.2 | |
| 38 | Kimi K2.7 Code | Moonshot AI | China | 2026-06-12 | yes | — | 68.4 | 26 | — | 150 (148 - 152) | — | — | — | — | 74.0 | 45.7 | 61 | 26 | 1 | 16.1 | |
| 39 | MiMo-V2.5-Pro | Xiaomi | China | 2026-04-22 | yes | 1467 ±4 | — | 26 | 53.01 | — | 40.97 ±1.38 | — | 1513 ±11 | 1521 ±6 | — | — | 45 | 30 | 0 | 9.4 | |
| 40 | Hunyuan Hy3 | Tencent | China | 2026-07-06 | yes | 1456 ±7 | — | 25 | 62.13 | — | — | — | 1520 ±25 | 1503 ±13 | — | — | 87 | 27 | 1 | — | |
| 41 | MiMo-V2.5 | Xiaomi | China | 2026-04-22 | yes | 1434 ±4 | — | 25* | — | — | 39.91 ±1.31 | — | 1483 ±13 | 1491 ±7 | — | — | 34 | 24 | 0 | — | |
| 42 | Qwen3.7-Plus | Alibaba | China | 2026-06-01 | yes | 1456 ±4 | — | 25 | — | 147 (146 - 149) | 38.65 ±1.17 | — | 1506 ±13 | 1501 ±7 | — | — | 65 | 13 | 1 | 9.4 | Released 2026-06-01 (10:00 China time) per the Qwen blog, the day Model Studio also lists it; Epoch records 2026-06-02, the China-time day of the launch post on X. |
| 43 | LongCat-2.0 | Meituan | China | 2026-06-30 | yes | — | — | 19 | 54.22 | — | — | — | — | — | — | — | — | 18 | 0 | — | Artificial Analysis no longer printed an output-speed value for this row on 2026-09-12, so the old figure was removed rather than retained without a live source. |
| 44 | Step 3.7 Flash | StepFun | China | 2026-05-29 | yes | — | — | 19* | 48.16 | — | — | — | — | — | — | — | 193 | 17 | — | — | |
| 45 | Step 3.5 Flash | StepFun | China | 2026-02-02 | yes | 1394 ±4 | — | 17* | — | — | — | — | 1435 ±11 | 1450 ±6 | — | — | 126 | — | — | — | StepFun's platform documentation lists step-3.5-flash as the base version and step-3.5-flash-2603 as an agent-optimised variant built from it; the 2603 build shipped on 2026-04-02 and has its own row here. LMArena added step-3.5-flash to its text board on 2026-02-10, linking the original open weights, seven weeks before 2603 shipped, so its three text-board rows are this build (57,137 votes on the overall board as of 2026-09-22). Artificial Analysis lists this build as Step 3.5 Flash (released 2026-02-02 on its model page) and marks the row deprecated; the leaderboard hides such rows under its default Status: Current filter, so its figures are read with the Status filter set to All, and its Intelligence Index is an estimate (*). The vendor's own SWE-bench Verified figure (74.4) that stood here while no independent board was quoted has been removed. LiveBench, Epoch, Vals and SuperCLUE carry no row for it. |
| 46 | Step 3.5 Flash 2603 | StepFun | China | 2026-04-02 | yes | — | — | 17* | — | — | — | — | — | — | — | — | 118 | — | — | — | StepFun's agent-optimised variant of step-3.5-flash, whose reasoning_effort takes only low and high. Released 2026-04-02 per StepFun's announcement on its X account ("Step 3.5 Flash 2603 just shipped!"). Artificial Analysis lists it as Step 3.5 Flash 2603 (released 2026-04-02 on its model page) and marks the row deprecated; the leaderboard hides such rows under its default Status: Current filter, so its figures are read with the Status filter set to All, and its Intelligence Index is an estimate (*). LMArena carries no 2603 row: its step-3.5-flash row is the base build and is not reused here. LiveBench, Epoch, Vals and SuperCLUE carry no row for it either. |
| 47 | GPT-5.5 | OpenAI | International | 2026-04-23 | yes | 1482 ±4 | 80.2 | — | — | 159 (157 - 162) | 57.41 ±1.15 | — | 1515 ±10 | 1520 ±6 | 82.1 | 54.0 | — | — | — | — | |
| 48 | Kimi K2.6 | Moonshot AI | China | 2026-04-20 | yes | 1460 ±5 | 70.5 | — | — | 151 (149 - 153) | 43.47 ±1.17 | — | 1526 ±14 | 1514 ±7 | 78.6 | 46.9 | — | — | — | — | |
| 49 | Qwen3.6-Plus | Alibaba | China | 2026-04-01 | yes | 1443 ±4 | 68.9 | — | — | 148 (145 - 150) | 31.98 ±0.98 | — | 1477 ±13 | 1495 ±6 | 78.2 | 41.4 | — | — | — | — | |
| 50 | Qwen3.7-Max | Alibaba | China | 2026-05-20 | yes | 1473 ±10 | 73.1 | — | — | 154 (152 - 156) | 44.77 ±1.08 | — | 1518 ±38 | 1525 ±18 | 74.2 | 43.6 | — | — | — | — | LMArena lists the preview build (qwen3.7-max-preview); the released model has no separate row. Released 2026-05-20 (10:00 China time) per the Qwen blog; Epoch records 2026-05-19, and the blog itself renders the date in the reader's time zone, so US readers see the 19th. |
| 51 | MiniMax-M2.7 | MiniMax | China | 2026-03-18 | yes | 1415 ±4 | — | — | — | 146 (138 - 148) | 24.56 ±0.93 | — | 1443 ±10 | 1481 ±5 | — | — | — | — | — | — | |
| 52 | GLM-5 | Zhipu AI (Z.ai) | China | 2026-02-12 | yes | 1458 ±4 | — | — | — | 146 (144 - 148) | — | — | 1515 ±15 | 1498 ±7 | — | — | — | — | — | — | Released 2026-02-12 per Z.ai's blog (GLM-5: From Vibe Coding to Agentic Engineering) and both of its release notes; Epoch records 2026-02-11, the US-time day of the launch post. |
| 53 | GLM-5.1 | Zhipu AI (Z.ai) | China | 2026-04-07 | yes | 1466 ±4 | — | — | — | 150 (148 - 152) | — | — | 1516 ±12 | 1513 ±6 | — | — | — | — | — | — | The Vals Index V2 (2026-08-13) no longer lists this model; the V1.2 figure it once carried was removed with that revision. |
| 54 | DeepSeek V4 Flash | DeepSeek | China | — | no | 1436 ±4 | — | — | — | — | — | — | 1472 ±12 | 1483 ±6 | — | — | — | — | — | — | The pre-0731 build. LMArena still lists it alongside 0731; LiveBench dropped its row between the 2026-08-19 and 2026-08-26 snapshots, so its 65.5 is no longer quoted. |
| 55 | GLM-5V-Turbo | Zhipu AI (Z.ai) | China | — | yes | 1433 ±7 | — | — | — | — | — | — | 1479 ±24 | 1489 ±12 | — | — | — | — | — | — | Vision-language model. LMArena's text board now carries a glm-5v-turbo row (9,361 votes as of 2026-09-22); Epoch does not score it. |
| 56 | Doubao-Seed-2.1-pro | ByteDance | China | 2026-06-28 | yes | — | — | — | 65.94 | — | — | Terminal Bench 2.1 71.0 | — | — | — | — | — | — | — | — | |
| 57 | Hy4-preview | Tencent | China | 2026-08-28 | yes | — | — | — | — | — | 55.40 ±1.25 | — | — | — | — | — | — | — | — | — | Tencent announced and open-sourced Hy4 preview on 2026-08-28. The Vals Index lists it as Hy4 Preview (tencent/hy4-preview); Artificial Analysis, LMArena, LiveBench and Epoch carried no row for it as of 2026-09-21, and those figures are left blank. Hy3 scores and AutoEval results are not attributed to this row. |
| 58 | Qwen3.6-Flash | Alibaba | China | 2026-04-26 | yes | — | — | — | — | 143 (141 - 145) | — | — | — | — | — | — | — | — | — | — | |
| 59 | Agnes 2.0 Flash | Agnes AI | China | — | no | — | — | — | — | — | — | Claw-Eval Pass^3 60.9% | — | — | — | — | — | — | — | — | No independent board in this snapshot covers it; the figure shown is the vendor's own. Retired from sale 2026-09-03 (CA-654): the vendor deprecated it in favour of agnes-2.5-flash; kept as a ranked row with on_chinaapi false so the roster count and maintained.at still describe the 2026-08-31 board read. |
| 60 | Agnes 2.5 Flash | Agnes AI | China | — | yes | — | — | — | — | — | — | — | — | — | — | — | — | — | — | — | Artificial Analysis carries a different build of this family (Agnes 2.5 Pro Alpha, 40) and the vendor publishes its own figures only as a chart image, so there is no number here that can be quoted. |
| 61 | Doubao-Seed-2.1-turbo | ByteDance | China | 2026-06-28 | yes | — | — | — | — | — | — | Terminal Bench 2.1 67.6 | — | — | — | — | — | — | — | — | No independent board in this snapshot covers it; the figure shown is the vendor's own. The vendor writes the name without the 260628 snapshot suffix we serve it under. |
| 62 | Doubao-Seed-Evolving | ByteDance | China | — | yes | — | — | — | — | — | — | — | — | — | — | — | — | — | — | — | Doubao publishes Seed Evolving as a rolling model under the stable doubao-seed-evolving API ID, so capabilities can change without a dated version suffix. No independent score could be verified for this exact model on the boards cited by this snapshot as of 2026-09-12; Seed 2.1 scores are not reused. |
| 63 | GLM-5-Turbo | Zhipu AI (Z.ai) | China | 2026-03-16 | yes | — | — | — | — | — | — | — | — | — | — | — | — | — | — | — | No independent board in this snapshot carries a row for it. |
| 64 | MiMo-V2.6-Flash | Xiaomi | China | 2026-09-22 | yes | — | — | — | — | — | — | DeepSWE v1.1 65.7 | — | — | — | — | — | — | — | — | Released and open-sourced 2026-09-22 per Xiaomi's release record (MiMo-V2.6 Series Release). No independent board in this snapshot covers it; the figure shown is the vendor's own. |
| 65 | MiMo-V2.6-Pro-Ultraspeed | Xiaomi | China | 2026-09-22 | yes | — | — | — | — | — | — | — | — | — | — | — | — | — | — | — | Xiaomi's high-speed serving mode of MiMo-V2.6-Pro, released 2026-09-22. The vendor states flagship V2.6-Pro performance at up to 20x the output speed but publishes no separate figures, and no board lists it separately, so MiMo-V2.6-Pro's scores are not reused here. |
Image generation models
Quality coverage: 13/17 models, 4 boards from 2 evaluation operators. Public preference evidence, not comprehensive capability certification or a ChinaAPI endpoint test. Tasks and audio pools are separate; scores across columns are not interchangeable. Overlapping intervals do not establish a significant lead.
Ordered by LMArena Text-to-Image Arena; a model with no figure on that board follows the ones that have one, because absent from a board is not the same as last on it.
Third-party price and generation-time references
Not ChinaAPI prices or endpoint measurements. These references never affect quality coverage or ranking. Compare only matching tasks and configurations. Generation time is a trailing-three-day median of successful requests including download, not retry overhead.
Video generation models
Quality coverage: 15/18 models, 4 boards from 2 evaluation operators. Public preference evidence, not comprehensive capability certification or a ChinaAPI endpoint test. Tasks and audio pools are separate; scores across columns are not interchangeable. Overlapping intervals do not establish a significant lead.
Ordered by LMArena Text-to-Video Arena; a model with no figure on that board follows the ones that have one, because absent from a board is not the same as last on it.
Third-party price and generation-time references
Not ChinaAPI prices or endpoint measurements. These references never affect quality coverage or ranking. Compare only matching tasks and configurations. Generation time is a trailing-three-day median of successful requests including download, not retry overhead.
Real usage on ChinaAPI
Which models the calls ChinaAPI actually served went to, as positions and shares of this platform's own traffic. It is a measurement of one gateway, not of the Chinese AI market, and it is not comparable with the benchmark results above. No call count, token count or revenue figure is published here or anywhere else on this site: every number below is a proportion. Recomputed hourly. A model with too little traffic in a window to rank is left unnamed and counted under “other”, as is anything outside the published catalogue. The same board as JSON is at /api/model-rankings/usage.
Share of calls served, the last 7 days
| # | Model | Vendor | Share of calls | Trend vs the previous 7 days |
|---|---|---|---|---|
| 1 | glm-5.3-flash | 智谱 | 5.9% | up |
| 2 | deepseek-flash | DeepSeek | 4.6% | down |
| 3 | hy3 | Tencent | 4.4% | up |
| 4 | claude-opus-4-8 | Anthropic | 2.9% | up |
| 5 | gpt-5.6-sol | OpenAI | 2.5% | up |
| 6 | MiniMax-M3 | MiniMax | 1.9% | flat |
| 7 | agnes-3.0-flash | Agnes AI | 1.7% | up |
| 8 | kimi-k2.6 | Moonshot | 1.7% | flat |
| 9 | claude-opus-5 | Anthropic | 1.5% | up |
| 10 | gpt-5.6-luna | OpenAI | 1.5% | up |
| 11 | claude-sonnet-5 | Anthropic | 1.4% | flat |
| 12 | deepseek-v4-pro | DeepSeek | 1.4% | down |
| 13 | gpt-5.6-terra | OpenAI | 1.3% | up |
| 14 | claude-opus-4-7 | Anthropic | 1.3% | flat |
| 15 | claude-fable-5 | Anthropic | 1.3% | flat |
| 16 | claude-sonnet-4-6 | Anthropic | 1.2% | flat |
| 17 | claude-opus-4-6 | Anthropic | 1.2% | flat |
| 18 | claude-haiku-4-5 | Anthropic | 1.2% | up |
| 19 | step-audio-2 | StepFun | 1.1% | flat |
| 20 | agnes-2.5-flash | Agnes AI | 1.0% | down |
| All other models | 58.9% |
How those calls arrived: 9.9% over agent protocols (Claude Messages and the OpenAI Responses API), 89.1% over chat protocols (OpenAI chat completions and Gemini generateContent), 1.0% over everything else (audio, images, embeddings, rerank and asynchronous tasks).
Share of calls served, the last 30 days
| # | Model | Vendor | Share of calls | Trend vs the previous 30 days |
|---|---|---|---|---|
| 1 | agnes-2.5-flash | Agnes AI | 4.6% | up |
| 2 | deepseek-flash | DeepSeek | 4.4% | up |
| 3 | deepseek-v4-pro | DeepSeek | 4.3% | down |
| 4 | glm-5.3-flash | 智谱 | 3.3% | up |
| 5 | gpt-5.6-sol | OpenAI | 2.1% | flat |
| 6 | kimi-k2.7-code | Moonshot | 2.0% | up |
| 7 | hy3 | Tencent | 1.8% | up |
| 8 | claude-opus-4-8 | Anthropic | 1.7% | flat |
| 9 | gpt-5.6-luna | OpenAI | 1.6% | flat |
| 10 | kimi-k2.6 | Moonshot | 1.5% | up |
| 11 | claude-opus-5 | Anthropic | 1.4% | flat |
| 12 | claude-sonnet-5 | Anthropic | 1.3% | down |
| 13 | MiniMax-M3 | MiniMax | 1.3% | up |
| 14 | kimi-k3 | Moonshot | 1.2% | down |
| 15 | claude-opus-4-7 | Anthropic | 1.2% | up |
| 16 | claude-sonnet-4-6 | Anthropic | 1.2% | up |
| 17 | gpt-5.6-terra | OpenAI | 1.2% | flat |
| 18 | claude-opus-4-6 | Anthropic | 1.2% | up |
| 19 | claude-fable-5 | Anthropic | 1.1% | flat |
| 20 | glm-5.2 | 智谱 | 1.1% | down |
| All other models | 60.3% |
How those calls arrived: 4.5% over agent protocols (Claude Messages and the OpenAI Responses API), 93.6% over chat protocols (OpenAI chat completions and Gemini generateContent), 1.9% over everything else (audio, images, embeddings, rerank and asynchronous tasks).