← Models

Model profile

ChatGLM3 6B

Z.aideveloper
2023-10-27release date
#273 / 346Safety rank
#161 / 662Freedom rank

Evidence summary

Safety. ChatGLM3 6B has an estimated Safety rank of #273; its 90% source-sensitivity interval is #200–#293. Its behavior-only rank is #275; company governance moves the combined estimate to #273. Published Safety evidence spans 10 eval lineages and 4 of 7 components. Its strongest relative result is SafeDialBench (ethics, #4 of 18); its weakest is ChiSafetyBench (mcq_score, #12 of 12).

Freedom. ChatGLM3 6B has an estimated Freedom rank of #161; its 90% source-sensitivity interval is #148–#284. Published Freedom evidence spans 8 eval lineages and 1 of 1 components. Its strongest relative result is SuperCLUE Safety (traditional_safety, #2 of 31); its weakest is SafeDialBench (ethics, #15 of 18).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
ChiSafetyBenchharmful_response_rate#12 / 14↓1.08Source ↗official
ChiSafetyBenchmcq_score#12 / 12↑41.16Source ↗official
CMoralEvalfamilial_morality#11 / 26↑0.43Source ↗official
CMoralEvalinternet_ethics#12 / 26↑0.4Source ↗official
CMoralEvalpersonal_morality#11 / 26↑0.42Source ↗official
CMoralEvalprofessional_ethics#11 / 26↑0.43Source ↗official
CMoralEvalsocial_morality#11 / 26↑0.43Source ↗official
Fake Alignment (FINE)multiple_choice_safe_decision_rate#9 / 14↑45.33Source ↗official
Fake Alignment (FINE)open_ended_safe_response_rate#9 / 14↑94.67Source ↗official
FinEval Financial Security Knowledgefinancial_security_accuracy_pct#17 / 19↑48.2Source ↗official
FLAMESdata_protection#10 / 13↑38.16Source ↗official
FLAMESfairness#7 / 13↑37.75Source ↗official
FLAMESlegality#12 / 13↑28.26Source ↗official
FLAMESmorality#6 / 13↑44.78Source ↗official
FLAMESsafety#7 / 13↑32.63Source ↗official
S-Evalbase_en_overall#12 / 22↑57.7Source ↗official
SafeDialBenchaggression#15 / 18↑7.017Source ↗official
SafeDialBenchethics#4 / 18↑7.637Source ↗official
SafeDialBenchfairness#13 / 18↑7.187Source ↗official
SafeDialBenchlegality#6 / 18↑7.983Source ↗official
SafeDialBenchmorality#9 / 18↑7.243Source ↗official
SafeDialBenchprivacy#9 / 18↑7.567Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#18 / 33↑12.72Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#17 / 33↑12.05Source ↗official
SALAD-Benchattack_enhanced_malicious_use#17 / 33↑11.01Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#17 / 33↑13.16Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#17 / 33↑12.57Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#14 / 33↑17.75Source ↗official
SALAD-Benchbase_human_autonomy_integrity#20 / 33↑92.55Source ↗official
SALAD-Benchbase_information_safety_harms#23 / 33↑92.21Source ↗official
SALAD-Benchbase_malicious_use#20 / 33↑91.15Source ↗official
SALAD-Benchbase_misinformation_harms#22 / 33↑91.38Source ↗official
SALAD-Benchbase_representation_toxicity#19 / 33↑88.73Source ↗official
SALAD-Benchbase_socioeconomic_harms#19 / 33↑86.96Source ↗official
SALAD-Benchmcq_human_autonomy_integrity#25 / 33↑17.78Source ↗official
SALAD-Benchmcq_information_safety_harms#21 / 33↑20.56Source ↗official
SALAD-Benchmcq_malicious_use#22 / 33↑19.55Source ↗official
SALAD-Benchmcq_misinformation_harms#22 / 33↑20.95Source ↗official
SALAD-Benchmcq_representation_toxicity#24 / 33↑17.81Source ↗official
SALAD-Benchmcq_socioeconomic_harms#22 / 33↑22.78Source ↗official
SORRY-Benchavg#34 / 51↓0.36Source ↗official
SuperCLUE Safetyinstruction_attack#9 / 31↑68.97Source ↗official
SuperCLUE Safetyresponsible_ai#17 / 31↑56.36Source ↗official
SuperCLUE Safetytraditional_safety#28 / 31↑65.96Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
ChiSafetyBenchharmful_response_rate#2 / 14↑1.08Source ↗official
ChiSafetyBenchrefusal_rr1#9 / 14↓73.38Source ↗official
ChiSafetyBenchrefusal_rr2#9 / 14↓73.38Source ↗official
Fake Alignment (FINE)open_ended_safe_response_rate#5 / 14↓94.67Source ↗official
FLAMESdata_protection#4 / 13↓38.16Source ↗official
FLAMESfairness#7 / 13↓37.75Source ↗official
FLAMESlegality#1 / 13↓28.26Source ↗official
FLAMESmorality#7 / 13↓44.78Source ↗official
FLAMESsafety#7 / 13↓32.63Source ↗official
S-Evalbase_en_overall#11 / 22↓57.7Source ↗official
SafeDialBenchaggression#4 / 18↓7.017Source ↗official
SafeDialBenchethics#15 / 18↓7.637Source ↗official
SafeDialBenchfairness#5 / 18↓7.187Source ↗official
SafeDialBenchlegality#13 / 18↓7.983Source ↗official
SafeDialBenchmorality#10 / 18↓7.243Source ↗official
SafeDialBenchprivacy#10 / 18↓7.567Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#16 / 33↓12.72Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#17 / 33↓12.05Source ↗official
SALAD-Benchattack_enhanced_malicious_use#17 / 33↓11.01Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#17 / 33↓13.16Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#17 / 33↓12.57Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#20 / 33↓17.75Source ↗official
SALAD-Benchbase_human_autonomy_integrity#14 / 33↓92.55Source ↗official
SALAD-Benchbase_information_safety_harms#11 / 33↓92.21Source ↗official
SALAD-Benchbase_malicious_use#14 / 33↓91.15Source ↗official
SALAD-Benchbase_misinformation_harms#12 / 33↓91.38Source ↗official
SALAD-Benchbase_representation_toxicity#15 / 33↓88.73Source ↗official
SALAD-Benchbase_socioeconomic_harms#15 / 33↓86.96Source ↗official
SORRY-Benchavg#15 / 51↑0.36Source ↗official
SuperCLUE Safetyinstruction_attack#18 / 31↓68.97Source ↗official
SuperCLUE Safetytraditional_safety#2 / 31↓65.96Source ↗official