← Models

Model profile

Baichuan 2 7B Chat

Baichuandeveloper
2023-09-06release date
#264 / 346Safety rank
#206 / 662Freedom rank

Evidence summary

Safety. Baichuan 2 7B Chat has an estimated Safety rank of #264; its 90% source-sensitivity interval is #183–#277. Its behavior-only rank is #267; company governance moves the combined estimate to #264. Published Safety evidence spans 11 eval lineages and 5 of 7 components. Its strongest relative result is FLAMES (safety, #1 of 13); its weakest is CRiskEval (improvement_intent, #17 of 17).

Freedom. Baichuan 2 7B Chat has an estimated Freedom rank of #206; its 90% source-sensitivity interval is #137–#328. Published Freedom evidence spans 7 eval lineages and 1 of 1 components. Its strongest relative result is DSPSafeBench (score, #2 of 12); its weakest is FLAMES (safety, #13 of 13).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
ChineseSafescore#13 / 22↑53.99Source ↗official
ChiSafetyBenchharmful_response_rate#8 / 14↓0.65Source ↗official
ChiSafetyBenchmcq_score#11 / 12↑50.03Source ↗official
CMoralEvalfamilial_morality#8 / 26↑0.47Source ↗official
CMoralEvalinternet_ethics#6 / 26↑0.48Source ↗official
CMoralEvalpersonal_morality#6 / 26↑0.47Source ↗official
CMoralEvalprofessional_ethics#6 / 26↑0.49Source ↗official
CMoralEvalsocial_morality#6 / 26↑0.5Source ↗official
CRiskEvaldeception_willingness#16 / 17↓38.4Source ↗official
CRiskEvaldesire_for_resource#14 / 17↓39.06Source ↗official
CRiskEvalharmful_goal#14 / 17↓52.04Source ↗official
CRiskEvalimprovement_intent#17 / 17↓60.08Source ↗official
CRiskEvalmalicious_coordination#16 / 17↓29.29Source ↗official
CRiskEvalself_preservation#13 / 17↓40.2Source ↗official
CRiskEvalsituational_awareness#6 / 17↓61.78Source ↗official
DSPSafeBenchscore#11 / 12↑65.31Source ↗official
Fake Alignment (FINE)multiple_choice_safe_decision_rate#12 / 14↑20Source ↗official
Fake Alignment (FINE)open_ended_safe_response_rate#5 / 14↑97.33Source ↗official
FLAMESdata_protection#8 / 13↑40.79Source ↗official
FLAMESfairness#4 / 13↑42.17Source ↗official
FLAMESlegality#4 / 13↑52.17Source ↗official
FLAMESmorality#11 / 13↑39.3Source ↗official
FLAMESsafety#1 / 13↑56.41Source ↗official
HarmBenchdr#13 / 28↓18.8Source ↗official
M3-SafetyBenchoverall_score#16 / 19↑82.91Source ↗official
SafeDialBenchaggression#11 / 18↑7.073Source ↗official
SafeDialBenchethics#6 / 18↑7.613Source ↗official
SafeDialBenchfairness#16 / 18↑7.123Source ↗official
SafeDialBenchlegality#9 / 18↑7.937Source ↗official
SafeDialBenchmorality#4 / 18↑7.383Source ↗official
SafeDialBenchprivacy#11 / 18↑7.523Source ↗official
SuperCLUE Safetyinstruction_attack#15 / 31↑65.52Source ↗official
SuperCLUE Safetyresponsible_ai#13 / 31↑61.82Source ↗official
SuperCLUE Safetytraditional_safety#5 / 31↑80.85Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
ChiSafetyBenchharmful_response_rate#5 / 14↑0.65Source ↗official
ChiSafetyBenchrefusal_rr1#7 / 14↓72.29Source ↗official
ChiSafetyBenchrefusal_rr2#7 / 14↓72.29Source ↗official
DSPSafeBenchscore#2 / 12↓65.31Source ↗official
Fake Alignment (FINE)open_ended_safe_response_rate#9 / 14↓97.33Source ↗official
FLAMESdata_protection#6 / 13↓40.79Source ↗official
FLAMESfairness#10 / 13↓42.17Source ↗official
FLAMESlegality#10 / 13↓52.17Source ↗official
FLAMESmorality#3 / 13↓39.3Source ↗official
FLAMESsafety#13 / 13↓56.41Source ↗official
HarmBenchdr#16 / 28↑18.8Source ↗official
SafeDialBenchaggression#8 / 18↓7.073Source ↗official
SafeDialBenchethics#13 / 18↓7.613Source ↗official
SafeDialBenchfairness#3 / 18↓7.123Source ↗official
SafeDialBenchlegality#10 / 18↓7.937Source ↗official
SafeDialBenchmorality#15 / 18↓7.383Source ↗official
SafeDialBenchprivacy#8 / 18↓7.523Source ↗official
SuperCLUE Safetyinstruction_attack#16 / 31↓65.52Source ↗official
SuperCLUE Safetytraditional_safety#23 / 31↓80.85Source ↗official