← Models

Model profile

Qwen1.5 7B Chat

Alibabadeveloper
2024-02-04release date
#220 / 346Safety rank
#273 / 662Freedom rank

Evidence summary

Safety. Qwen1.5 7B Chat has an estimated Safety rank of #220; its 90% source-sensitivity interval is #141–#257. Its behavior-only rank is #213; company governance moves the combined estimate to #220. Published Safety evidence spans 9 eval lineages and 6 of 7 components. Its strongest relative result is CRiskEval (malicious_coordination, #1 of 17); its weakest is SALAD-Bench (attack_enhanced_socioeconomic_harms, #27 of 33).

Freedom. Qwen1.5 7B Chat has an estimated Freedom rank of #273; its 90% source-sensitivity interval is #187–#396. Published Freedom evidence spans 6 eval lineages and 1 of 1 components. Its strongest relative result is SALAD-Bench (attack_enhanced_socioeconomic_harms, #6 of 33); its weakest is HyperCLOVA X Toxic Continuation Panels (kold_toxicity, #6 of 7).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
ChineseSafescore#6 / 22↑62.48Source ↗official
ChiSafetyBenchharmful_response_rate#5 / 14↓0.43Source ↗official
ChiSafetyBenchmcq_score#8 / 12↑79.39Source ↗official
Contextual MoralChoicehuman_agreement#12 / 22↑0.4Source ↗official
CRiskEvaldeception_willingness#9 / 17↓22.6Source ↗official
CRiskEvaldesire_for_resource#8 / 17↓32.14Source ↗official
CRiskEvalharmful_goal#9 / 17↓37.92Source ↗official
CRiskEvalimprovement_intent#6 / 17↓47.47Source ↗official
CRiskEvalmalicious_coordination#1 / 17↓5.68Source ↗official
CRiskEvalself_preservation#7 / 17↓36.01Source ↗official
CRiskEvalsituational_awareness#13 / 17↓67.2Source ↗official
HyperCLOVA X Toxic Continuation Panelskold_toxic_count#2 / 7↓0.0036Source ↗official
HyperCLOVA X Toxic Continuation Panelskold_toxicity#2 / 7↓0.1061Source ↗official
HyperCLOVA X Toxic Continuation Panelsrealt-toxicprompts_toxic_count#3 / 7↓0.006Source ↗official
HyperCLOVA X Toxic Continuation Panelsrealt-toxicprompts_toxicity#3 / 7↓0.055Source ↗official
JailBenchjailbreak_success_rate#11 / 14↓71.6Source ↗official
OR-Benchover_refusal_rate#10 / 25↓39.2Source ↗official
OR-Benchtoxic_acceptance_rate#18 / 25↓15Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#22 / 33↑9.91Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#21 / 33↑7.17Source ↗official
SALAD-Benchattack_enhanced_malicious_use#23 / 33↑8.32Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#19 / 33↑9.38Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#23 / 33↑7.39Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#27 / 33↑7.36Source ↗official
SALAD-Benchbase_human_autonomy_integrity#16 / 33↑95.17Source ↗official
SALAD-Benchbase_information_safety_harms#14 / 33↑95.67Source ↗official
SALAD-Benchbase_malicious_use#16 / 33↑94.12Source ↗official
SALAD-Benchbase_misinformation_harms#14 / 33↑94.04Source ↗official
SALAD-Benchbase_representation_toxicity#15 / 33↑90.23Source ↗official
SALAD-Benchbase_socioeconomic_harms#13 / 33↑91.19Source ↗official
SALAD-Benchmcq_human_autonomy_integrity#24 / 33↑19.17Source ↗official
SALAD-Benchmcq_information_safety_harms#26 / 33↑11.94Source ↗official
SALAD-Benchmcq_malicious_use#24 / 33↑16.35Source ↗official
SALAD-Benchmcq_misinformation_harms#24 / 33↑18.81Source ↗official
SALAD-Benchmcq_representation_toxicity#25 / 33↑15.62Source ↗official
SALAD-Benchmcq_socioeconomic_harms#24 / 33↑17.78Source ↗official
SORRY-Benchavg#39 / 51↓0.4Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
ChiSafetyBenchharmful_response_rate#8 / 14↑0.43Source ↗official
ChiSafetyBenchrefusal_rr1#10 / 14↓73.59Source ↗official
ChiSafetyBenchrefusal_rr2#10 / 14↓73.59Source ↗official
HyperCLOVA X Toxic Continuation Panelskold_toxic_count#6 / 7↑0.0036Source ↗official
HyperCLOVA X Toxic Continuation Panelskold_toxicity#6 / 7↑0.1061Source ↗official
HyperCLOVA X Toxic Continuation Panelsrealt-toxicprompts_toxic_count#5 / 7↑0.006Source ↗official
HyperCLOVA X Toxic Continuation Panelsrealt-toxicprompts_toxicity#5 / 7↑0.055Source ↗official
JailBenchjailbreak_success_rate#4 / 14↑71.6Source ↗official
OR-Benchover_refusal_rate#10 / 25↓39.2Source ↗official
OR-Benchtoxic_acceptance_rate#8 / 25↑15Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#11 / 33↓9.91Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#12 / 33↓7.17Source ↗official
SALAD-Benchattack_enhanced_malicious_use#11 / 33↓8.32Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#14 / 33↓9.38Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#11 / 33↓7.39Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#6 / 33↓7.36Source ↗official
SALAD-Benchbase_human_autonomy_integrity#18 / 33↓95.17Source ↗official
SALAD-Benchbase_information_safety_harms#20 / 33↓95.67Source ↗official
SALAD-Benchbase_malicious_use#18 / 33↓94.12Source ↗official
SALAD-Benchbase_misinformation_harms#19 / 33↓94.04Source ↗official
SALAD-Benchbase_representation_toxicity#19 / 33↓90.23Source ↗official
SALAD-Benchbase_socioeconomic_harms#20 / 33↓91.19Source ↗official
SORRY-Benchavg#13 / 51↑0.4Source ↗official