← Models

Model profile

Baichuan 2 13B Chat

Baichuandeveloper
2023-09-06release date
#186 / 346Safety rank
#548 / 662Freedom rank

Evidence summary

Safety. Baichuan 2 13B Chat has an estimated Safety rank of #186; its 90% source-sensitivity interval is #102–#239. Its behavior-only rank is #190; company governance moves the combined estimate to #186. Published Safety evidence spans 13 eval lineages and 5 of 7 components. Its strongest relative result is SuperCLUE Safety (traditional_safety, #1 of 31); its weakest is SafetyBench (UB, #20 of 21).

Freedom. Baichuan 2 13B Chat has an estimated Freedom rank of #548; its 90% source-sensitivity interval is #303–#637. Published Freedom evidence spans 7 eval lineages and 1 of 1 components. Its strongest relative result is SafeDialBench (fairness, #4 of 18); its weakest is SuperCLUE Safety (traditional_safety, #30 of 31).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
ChineseSafescore#1 / 22↑70.43Source ↗official
ChiSafetyBenchharmful_response_rate#5 / 14↓0.43Source ↗official
ChiSafetyBenchmcq_score#7 / 12↑79.83Source ↗official
CMoralEvalfamilial_morality#13 / 26↑0.39Source ↗official
CMoralEvalinternet_ethics#19 / 26↑0.34Source ↗official
CMoralEvalpersonal_morality#17 / 26↑0.36Source ↗official
CMoralEvalprofessional_ethics#19 / 26↑0.36Source ↗official
CMoralEvalsocial_morality#16 / 26↑0.37Source ↗official
CRiskEvaldeception_willingness#15 / 17↓37.75Source ↗official
CRiskEvaldesire_for_resource#11 / 17↓35.35Source ↗official
CRiskEvalharmful_goal#15 / 17↓54.94Source ↗official
CRiskEvalimprovement_intent#10 / 17↓50.82Source ↗official
CRiskEvalmalicious_coordination#12 / 17↓18.46Source ↗official
CRiskEvalself_preservation#15 / 17↓43.45Source ↗official
CRiskEvalsituational_awareness#7 / 17↓63.34Source ↗official
Fake Alignment (FINE)multiple_choice_safe_decision_rate#9 / 14↑45.33Source ↗official
Fake Alignment (FINE)open_ended_safe_response_rate#1 / 14↑100Source ↗official
FinEval Financial Security Knowledgefinancial_security_accuracy_pct#16 / 19↑61.6Source ↗official
FLAMESdata_protection#9 / 13↑39.47Source ↗official
FLAMESfairness#6 / 13↑38.55Source ↗official
FLAMESlegality#7 / 13↑39.13Source ↗official
FLAMESmorality#6 / 13↑44.78Source ↗official
FLAMESsafety#2 / 13↑53.85Source ↗official
HarmBenchdr#14 / 28↓19.3Source ↗official
M3-SafetyBenchoverall_score#13 / 19↑86.39Source ↗official
S-Evalbase_en_overall#4 / 22↑77.4Source ↗official
SafeDialBenchaggression#12 / 18↑7.03Source ↗official
SafeDialBenchethics#7 / 18↑7.6Source ↗official
SafeDialBenchfairness#15 / 18↑7.17Source ↗official
SafeDialBenchlegality#4 / 18↑7.987Source ↗official
SafeDialBenchmorality#6 / 18↑7.303Source ↗official
SafeDialBenchprivacy#5 / 18↑7.617Source ↗official
SafetyBenchEM#5 / 21↑75.75Source ↗official
SafetyBenchIA#4 / 21↑82.65Source ↗official
SafetyBenchMH#7 / 21↑83.65Source ↗official
SafetyBenchOFF#6 / 21↑69.25Source ↗official
SafetyBenchPH#5 / 21↑76.35Source ↗official
SafetyBenchPP#4 / 21↑82.05Source ↗official
SafetyBenchUB#20 / 21↑49.2Source ↗official
SuperCLUE Safetyinstruction_attack#3 / 31↑72.41Source ↗official
SuperCLUE Safetyresponsible_ai#4 / 31↑72.73Source ↗official
SuperCLUE Safetytraditional_safety#1 / 31↑87.23Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
ChiSafetyBenchharmful_response_rate#8 / 14↑0.43Source ↗official
ChiSafetyBenchrefusal_rr1#13 / 14↓77.06Source ↗official
ChiSafetyBenchrefusal_rr2#13 / 14↓76.84Source ↗official
Fake Alignment (FINE)open_ended_safe_response_rate#13 / 14↓100Source ↗official
FLAMESdata_protection#5 / 13↓39.47Source ↗official
FLAMESfairness#8 / 13↓38.55Source ↗official
FLAMESlegality#6 / 13↓39.13Source ↗official
FLAMESmorality#7 / 13↓44.78Source ↗official
FLAMESsafety#12 / 13↓53.85Source ↗official
HarmBenchdr#15 / 28↑19.3Source ↗official
S-Evalbase_en_overall#19 / 22↓77.4Source ↗official
SafeDialBenchaggression#7 / 18↓7.03Source ↗official
SafeDialBenchethics#12 / 18↓7.6Source ↗official
SafeDialBenchfairness#4 / 18↓7.17Source ↗official
SafeDialBenchlegality#14 / 18↓7.987Source ↗official
SafeDialBenchmorality#13 / 18↓7.303Source ↗official
SafeDialBenchprivacy#14 / 18↓7.617Source ↗official
SuperCLUE Safetyinstruction_attack#25 / 31↓72.41Source ↗official
SuperCLUE Safetytraditional_safety#30 / 31↓87.23Source ↗official