Chinese
中文Your shortlist. Side by side.
Accuracy by benchmark
Each benchmark measures a different set of questions.
Qwen3.7 Flash · High reasoningQwen3.7 Flash · Low reasoningSolar Pro 4 · Low reasoning
Accuracy · 0–100%
The same models, in other languages
Follow a column to see how a model’s accuracy changes with language.
| Language | Qwen3.7 FlashHigh reasoning | Qwen3.7 FlashLow reasoning | Solar Pro 4Low reasoning |
|---|---|---|---|
| ChineseThis language | 100.0% | 100.0% | 100.0% |
| Albanian | — | — | — |
| Czech | — | — | — |
| English | 96.0% | 96.0% | 96.0% |
| French | — | — | — |
| German | — | — | — |
| Japanese | — | — | — |
| Kazakh | 92.0% | 84.0% | 88.0% |
| Korean | — | — | — |
| Russian | 96.0% | 96.0% | 92.0% |
| Serbian | — | — | — |
| Slovak | — | — | — |
| Spanish | 92.0% | 92.0% | 80.0% |
| Swedish | — | — | — |
Chinese · Accuracy is the percentage of correct answers. Each available benchmark has equal weight in the overall score.