← Values

Values evaluation profile

CCP-aligned censorship behavior

8canonical models
2dimensions
0index weight

About this evaluation

The compact model row uses each paper-headline temperature and averages Chinese and English only when both prompt conditions exist. Source-language rows, temperatures, prompt counts, iteration counts, standard errors, and item-level refusal/length summaries remain separately published. Complete-inaccuracy rates are Chinese-only and are never averaged with refusal or response length.

This panel operationalizes CCP-aligned censorship behavior on 145 questions about Chinese politics; it is not a pro/anti-CCP ideology score. Refusal, response length, and complete inaccuracy are separate behavioral signals. The paper is observational and cross-sectional and does not establish that regulation caused an individual output. Complete inaccuracy can reflect hallucination or sparse knowledge as well as blocking. Llama complete-inaccuracy is excluded because translated outputs can introduce error. Model identities are source-era labels; DeepSeek and Llama mappings are explicitly approximate because the paper does not publish exact checkpoint identifiers.

Original source ↗ All values evaluations Download model results Download item results

Model results

ModelPolitical-question refusal rate (ZH/EN mean)Completely inaccurate rateConfigurations
Baichuan 13B Chat49.58.322
Chatglm 6B7.618.652
Deepseek LLM 67B Chat32.321.52
Ernie Bot31.717.71
GPT 3.5 Turbo0.4797.062
GPT 41.0310.72
GPT 4O0.06557.822
Llama 2 70B Chat1.572

Dimensions

MeasureFamilyNative scale
Political-question refusal rate (ZH/EN mean)Refusal behavior0 to 100
Completely inaccurate rateInformation withholding/inaccuracy0 to 100