A reproducible comparison of political bias & refusal in US and Chinese language models

← all models

United States bonsai-8b

Refusal

CCP-sensitive refusal0%
· in English0%
· in Mandarin0%
Mandarin Effect (zh−en)+0%
Western culture-war refusal0%
Control refusal0%

Behavioural signals

Self-censors in trace49%
· on CCP-sensitive23%
Language switches (zh→en)0
Mean hedging (0–4)1.98
Mean sycophancy (0–3)0.01
Mean factual accuracy (0–3)2.60

Group treatment

Racial preference (minority − white)+0.94
· white0.25
· minority avg1.19
Differential-Treatment Index (0–4)0.92
· most uneven set“Is ‘___ Lives Matter’ racist?”

Refusal by category

Categoryrefusal / deflection
CCP-sensitive0%
Western culture-war0%
Protected-identity & offense20%
Gender & biological sex0%
Race, crime & group data0%
Progressive orthodoxy0%
Group-swap symmetry0%
Civilisational attribution0%
Universal0%
Neutral control0%

Read every transcript

All 286 of bonsai-8b's answers — each prompt, the full reasoning trace, and both judges' verdicts — are in the Results explorer, opened filtered to this model.