AI Tools

Chinese Foundation Models

Every Chinese Foundation Models tool in the catalog, with its vendor, home jurisdiction, and the governance question it raises.

7 tools · 7 vendors · 1 jurisdictions

In short

Chinese Foundation Models covers 7 tools from 7 vendors across 1 jurisdictions. Each links to the vendor's own product page. Live model rankings from arena.ai (formerly LMSYS Chatbot Arena) sit below, refreshed from the source rather than typed by hand. This is a reference list to inventory against, not a recommendation to buy.

Current model rankings

Rankings below come from arena.ai (formerly LMSYS Chatbot Arena), fetched automatically and stamped with the date they were true. Model leaderboards move weekly, so any figure typed into a page by hand is wrong within days. These are not.

As of 2026-08-13 Source: arena.ai (formerly LMSYS Chatbot Arena)

Read this before reading the tables

Crowdsourced blind pairwise voting, scored with a Bradley-Terry model and reported as an Elo-style rating with a 95 percent confidence interval. Gaps under roughly 10 points sit inside the noise floor and should not be read as a ranking.

Overall text and chat

Head-to-head human preference on general conversation. The closest thing the field has to a general-purpose ranking.

# Model Vendor Rating ± 95% CI
1 claude-fable-5 Anthropic 1,507 ± 5
2 claude-opus-4-6-high Anthropic 1,505 ± 4
3 claude-opus-4-7-high Anthropic 1,502 ± 4
4 muse-spark-1.2 (xHigh) Meta 1,499 ± 10
5 claude-opus-4-6 Anthropic 1,497 ± 3
6 claude-opus-4-7 Anthropic 1,494 ± 4
7 claude-opus-5-high Anthropic 1,494 ± 5
8 claude-opus-5-max Anthropic 1,491 ± 7
9 qwen3.8-max Alibaba 1,491 ± 8
10 muse-spark-1.1 Meta 1,489 ± 6
11 kimi-k3-max Moonshot 1,489 ± 6
12 muse-spark Meta 1,488 ± 6

Top 12 of 30 ranked. Full board at arena.ai (formerly LMSYS Chatbot Arena) →

Code generation

The same vote mechanic restricted to coding prompts. Ranks differ sharply from the text board, which is the reason to read both.

# Model Vendor Rating ± 95% CI
1 claude-opus-5-max Anthropic 1,691 ± 10
2 kimi-k3-max Moonshot 1,674 ± 11
3 qwen3.8-max Alibaba 1,669 ± 14
4 claude-opus-5-high Anthropic 1,664 ± 9
5 claude-fable-5 Anthropic 1,627 ± 9
6 gpt-5.6-sol-xhigh (codex-harness) OpenAI 1,622 ± 8
7 grok-4.6-high SpaceXAI 1,618 ± 21
8 glm-5.2-max Z.ai 1,587 ± 8
9 deepseek-v4-flash-high DeepSeek 1,582 ± 12
10 claude-opus-4-8-high Anthropic 1,564 ± 7
11 claude-opus-4-7 Anthropic 1,558 ± 6
12 claude-opus-4-7-high Anthropic 1,557 ± 6

Top 12 of 50 ranked. Full board at arena.ai (formerly LMSYS Chatbot Arena) →

A leaderboard measures average preference across a crowd of prompts that are not your prompts. It says nothing about your data terms, your jurisdiction, your latency budget, or your integration surface, and those usually decide procurement. Use it to shortlist, then evaluate on your own workload. Benchmarks beyond these boards, including MMLU-Pro, SWE-bench Verified, GPQA Diamond, and throughput figures, are deliberately not reproduced here: no verified machine-readable source for them is wired up yet, and an unsourced number on this site would be worse than a missing one.

All 7 Chinese Foundation Models tools

Frequently asked questions

How many Chinese Foundation Models tools are there?

This catalog lists 7 tools in the Chinese Foundation Models category, from 7 vendors across 1 jurisdictions. 0 carry a full SRJ profile; the rest link to the vendor's own product page.

Which model ranks highest right now?

On the overall text and chat board at arena.ai (formerly LMSYS Chatbot Arena), as of 2026-08-13, the leader is claude-fable-5 from Anthropic on 1507 Elo. Rankings move weekly, so the table on this page is refreshed from the source automatically rather than typed by hand. Read the confidence intervals before treating a small gap as a difference.

Do benchmark rankings tell you which model to buy?

No. A leaderboard measures average preference across a crowd of prompts that are not your prompts. It says nothing about your data terms, your jurisdiction, your latency budget, or your integration surface, and those usually decide procurement. Use the rankings to shortlist, then evaluate on your own workload against your own governance requirements.

Does inclusion in this category mean SRJ recommends the tool?

No. This is a reference catalog, not an endorsement list. Presence here means only that a tool is common enough in the field that it belongs on an inventory checklist. It is not a security review, a due-diligence result, or a recommendation to buy.

Why does the catalog list each vendor's home jurisdiction?

Jurisdiction is a governance fact. Where a vendor is based determines which data-transfer, privacy, and cross-border rules apply before any business data reaches the tool. Two tools that look identical in features can carry very different obligations once the vendor's home country is taken into account. This category spans 1 jurisdictions.

Other categories

← Back to the full catalog

You cannot govern what you have not listed

The AI Business Enablement Audit™ builds the inventory, measures your organization against every framework in the AI Governance Reference Library, and delivers a defensible governance dossier.

Start or finish your AI Audit →