← Models

Model profile

Yi 6B Chat

01.AIdeveloper
2023-11-23release date
#172 / 267overall rank
5eval lineages

Evidence summary

Yi 6B Chat has an estimated overall rank of #172; its 90% source-sensitivity interval is #89–#203. Its behavior-only rank is #175; company governance moves the combined estimate to #172. Published evidence spans 5 evals and 5 of 7 behavior components. Its strongest relative result is CMoralEval (familial_morality, #4 of 26); its weakest is CRiskEval (self_preservation, #17 of 17).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
ChiSafetyBenchharmful_response_rate#11 / 140.87↓ lowerSource ↗official
ChiSafetyBenchmcq_score#5 / 1286.01↑ higherSource ↗official
CMoralEvalfamilial_morality#4 / 260.52↑ higherSource ↗official
CMoralEvalinternet_ethics#5 / 260.5↑ higherSource ↗official
CMoralEvalpersonal_morality#4 / 260.5↑ higherSource ↗official
CMoralEvalprofessional_ethics#5 / 260.5↑ higherSource ↗official
CMoralEvalsocial_morality#4 / 260.51↑ higherSource ↗official
CRiskEvaldeception_willingness#12 / 1726.92↓ lowerSource ↗official
CRiskEvaldesire_for_resource#15 / 1740.53↓ lowerSource ↗official
CRiskEvalharmful_goal#12 / 1742.13↓ lowerSource ↗official
CRiskEvalimprovement_intent#12 / 1752.06↓ lowerSource ↗official
CRiskEvalmalicious_coordination#14 / 1723.85↓ lowerSource ↗official
CRiskEvalself_preservation#17 / 1745.21↓ lowerSource ↗official
CRiskEvalsituational_awareness#5 / 1761.29↓ lowerSource ↗official
SafeDialBenchaggression#8 / 187.127↑ higherSource ↗official
SafeDialBenchethics#9 / 187.577↑ higherSource ↗official
SafeDialBenchfairness#10 / 187.277↑ higherSource ↗official
SafeDialBenchlegality#11 / 187.887↑ higherSource ↗official
SafeDialBenchmorality#15 / 187.123↑ higherSource ↗official
SafeDialBenchprivacy#4 / 187.67↑ higherSource ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#10 / 3323.71↑ higherSource ↗official
SALAD-Benchattack_enhanced_information_safety_harms#8 / 3327.04↑ higherSource ↗official
SALAD-Benchattack_enhanced_malicious_use#9 / 3323.82↑ higherSource ↗official
SALAD-Benchattack_enhanced_misinformation_harms#10 / 3324.84↑ higherSource ↗official
SALAD-Benchattack_enhanced_representation_toxicity#12 / 3320.24↑ higherSource ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#10 / 3323.38↑ higherSource ↗official
SALAD-Benchbase_human_autonomy_integrity#25 / 3386.78↑ higherSource ↗official
SALAD-Benchbase_information_safety_harms#20 / 3393.77↑ higherSource ↗official
SALAD-Benchbase_malicious_use#26 / 3382.67↑ higherSource ↗official
SALAD-Benchbase_misinformation_harms#27 / 3386.12↑ higherSource ↗official
SALAD-Benchbase_representation_toxicity#27 / 3378.37↑ higherSource ↗official
SALAD-Benchbase_socioeconomic_harms#20 / 3386.72↑ higherSource ↗official
SALAD-Benchmcq_human_autonomy_integrity#28 / 335.833↑ higherSource ↗official
SALAD-Benchmcq_information_safety_harms#28 / 333.889↑ higherSource ↗official
SALAD-Benchmcq_malicious_use#27 / 335.962↑ higherSource ↗official
SALAD-Benchmcq_misinformation_harms#28 / 335.238↑ higherSource ↗official
SALAD-Benchmcq_representation_toxicity#27 / 334.792↑ higherSource ↗official
SALAD-Benchmcq_socioeconomic_harms#28 / 336.667↑ higherSource ↗official