← Values

Values evaluation profile

Taiwan Sovereignty Benchmark Pro

16canonical models
4dimensions
0index weight

About this evaluation

Mean Chinese/English prompt pass rates are shown as a compact headline because the prompts are paired translations scored by the same keyword rubric. Chinese and English rates, consistency, warning rates, soft-censorship match rates, API-error rates, and language gaps remain separately published. Repeated runs are not averaged: the all-error GPT-4o Mini run is excluded and retained in the audit artifact.

This is compatibility with an explicitly pro-Taiwan, source-authored keyword rubric—not geopolitical truth, neutrality, democracy quality, or generic model safety. Neutral quotation or discussion of PRC claims can trigger a fail; expected factual elements are not scored; soft-censorship phrases do not themselves fail; most cells are one sample from mutable hosted endpoints. Qwen3 Max API failures are preserved separately and its pass rate must not be read as pure narrative agreement.

Original source ↗ All values evaluations Download model results Download language results Download item results

Model results

ModelPro-Taiwan rubric compatibilityWarning-phrase rateSoft-censorship rateAPI-error rateConfigurations
Claude 3.5 Sonnet900001
Claude Opus 4.5900501
Claude Sonnet 4.5700001
Deepseek V31040001
Gemini 2.0 Flash655001
Gemini 3 Pro Preview6015001
GPT 4O900001
GPT 4O Mini100.00001
GPT 5.28510001
Grok 355201001
Kimi K2.51540001
Llama 3.3 70B Instruct900001
Minimax M25535001
Mistral Large 34530501
Qwen2.5 72B Instruct2010501
Qwen3 Max0200351

Dimensions

MeasureFamilyNative scale
Pro-Taiwan rubric compatibilityTaiwan sovereignty framing0 to 100
Warning-phrase rateResponse behavior0 to 100
Soft-censorship rateResponse behavior0 to 100
API-error rateResponse behavior0 to 100