← Models

Model profile

Yi 34B Chat

01.AIdeveloper
2023-11-23release date
#207 / 346Safety rank
#296 / 662Freedom rank

Evidence summary

Safety. Yi 34B Chat has an estimated Safety rank of #207; its 90% source-sensitivity interval is #106–#254. Its behavior-only rank is #208; company governance moves the combined estimate to #207. Published Safety evidence spans 9 eval lineages and 5 of 7 components. Its strongest relative result is CMoralEval (familial_morality, #1 of 26); its weakest is COMPL-AI AI-Identity Disclosure (score, #14 of 14).

Freedom. Yi 34B Chat has an estimated Freedom rank of #296; its 90% source-sensitivity interval is #148–#494. Published Freedom evidence spans 7 eval lineages and 1 of 1 components. Its strongest relative result is S-Eval (base_en_overall, #3 of 22); its weakest is SafeDialBench (legality, #18 of 18).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AIRBench 2024 Safety Scenariossafety_scenarios#60 / 80↑0.536Source ↗official
ChiSafetyBenchharmful_response_rate#8 / 14↓0.65Source ↗official
ChiSafetyBenchmcq_score#10 / 12↑68.54Source ↗official
CMoralEvalfamilial_morality#1 / 26↑0.71Source ↗official
CMoralEvalinternet_ethics#1 / 26↑0.69Source ↗official
CMoralEvalpersonal_morality#1 / 26↑0.66Source ↗official
CMoralEvalprofessional_ethics#1 / 26↑0.7Source ↗official
CMoralEvalsocial_morality#1 / 26↑0.71Source ↗official
COMPL-AI AI-Identity Disclosurescore#14 / 14↑0.3562Source ↗official
COMPL-AI LLM RuLES Multi-Turn Rule Followingscore#4 / 14↑0.5829Source ↗official
COMPL-AI TensorTrust Goal-Hijacking Resistancescore#3 / 13↑0.5387Source ↗official
CRiskEvaldeception_willingness#3 / 17↓18.68Source ↗official
CRiskEvaldesire_for_resource#9 / 17↓32.56Source ↗official
CRiskEvalharmful_goal#10 / 17↓41.77Source ↗official
CRiskEvalimprovement_intent#11 / 17↓51.23Source ↗official
CRiskEvalmalicious_coordination#8 / 17↓10.33Source ↗official
CRiskEvalself_preservation#12 / 17↓39.99Source ↗official
CRiskEvalsituational_awareness#14 / 17↓67.23Source ↗official
S-Evalbase_en_overall#20 / 22↑39.3Source ↗official
SafeDialBenchaggression#3 / 18↑7.26Source ↗official
SafeDialBenchethics#2 / 18↑7.68Source ↗official
SafeDialBenchfairness#9 / 18↑7.337Source ↗official
SafeDialBenchlegality#1 / 18↑8.117Source ↗official
SafeDialBenchmorality#1 / 18↑7.52Source ↗official
SafeDialBenchprivacy#1 / 18↑7.88Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#9 / 33↑24.14Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#7 / 33↑27.36Source ↗official
SALAD-Benchattack_enhanced_malicious_use#11 / 33↑22.76Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#9 / 33↑26.81Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#10 / 33↑22.6Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#8 / 33↑23.81Source ↗official
SALAD-Benchbase_human_autonomy_integrity#21 / 33↑91.73Source ↗official
SALAD-Benchbase_information_safety_harms#21 / 33↑93.23Source ↗official
SALAD-Benchbase_malicious_use#21 / 33↑89.36Source ↗official
SALAD-Benchbase_misinformation_harms#26 / 33↑87.74Source ↗official
SALAD-Benchbase_representation_toxicity#26 / 33↑81.07Source ↗official
SALAD-Benchbase_socioeconomic_harms#17 / 33↑89.19Source ↗official
SALAD-Benchmcq_human_autonomy_integrity#21 / 33↑25Source ↗official
SALAD-Benchmcq_information_safety_harms#18 / 33↑31.67Source ↗official
SALAD-Benchmcq_malicious_use#18 / 33↑27.76Source ↗official
SALAD-Benchmcq_misinformation_harms#20 / 33↑26.43Source ↗official
SALAD-Benchmcq_representation_toxicity#20 / 33↑26.98Source ↗official
SALAD-Benchmcq_socioeconomic_harms#18 / 33↑31.67Source ↗official
SuperCLUE Safetyinstruction_attack#3 / 31↑72.41Source ↗official
SuperCLUE Safetyresponsible_ai#7 / 31↑69.09Source ↗official
SuperCLUE Safetytraditional_safety#10 / 31↑78.72Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AIRBench 2024 Safety Scenariossafety_scenarios#21 / 80↓0.536Source ↗official
ChiSafetyBenchharmful_response_rate#5 / 14↑0.65Source ↗official
ChiSafetyBenchrefusal_rr1#3 / 14↓69.7Source ↗official
ChiSafetyBenchrefusal_rr2#3 / 14↓69.7Source ↗official
COMPL-AI LLM RuLES Multi-Turn Rule Followingscore#11 / 14↓0.5829Source ↗official
S-Evalbase_en_overall#3 / 22↓39.3Source ↗official
SafeDialBenchaggression#16 / 18↓7.26Source ↗official
SafeDialBenchethics#16 / 18↓7.68Source ↗official
SafeDialBenchfairness#10 / 18↓7.337Source ↗official
SafeDialBenchlegality#18 / 18↓8.117Source ↗official
SafeDialBenchmorality#18 / 18↓7.52Source ↗official
SafeDialBenchprivacy#18 / 18↓7.88Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#25 / 33↓24.14Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#27 / 33↓27.36Source ↗official
SALAD-Benchattack_enhanced_malicious_use#23 / 33↓22.76Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#25 / 33↓26.81Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#24 / 33↓22.6Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#25 / 33↓23.81Source ↗official
SALAD-Benchbase_human_autonomy_integrity#13 / 33↓91.73Source ↗official
SALAD-Benchbase_information_safety_harms#13 / 33↓93.23Source ↗official
SALAD-Benchbase_malicious_use#13 / 33↓89.36Source ↗official
SALAD-Benchbase_misinformation_harms#8 / 33↓87.74Source ↗official
SALAD-Benchbase_representation_toxicity#8 / 33↓81.07Source ↗official
SALAD-Benchbase_socioeconomic_harms#17 / 33↓89.19Source ↗official
SuperCLUE Safetyinstruction_attack#25 / 31↓72.41Source ↗official
SuperCLUE Safetytraditional_safety#19 / 31↓78.72Source ↗official