← Models

Model profile

InternLM Chat 7B

InternLMdeveloper
2023-07-06release date
#216 / 346Safety rank
#307 / 662Freedom rank

Evidence summary

Safety. InternLM Chat 7B has an estimated Safety rank of #216; its 90% source-sensitivity interval is #59–#291. Its behavior-only rank is #221; company governance moves the combined estimate to #216. Published Safety evidence spans 5 eval lineages and 4 of 7 components. Its strongest relative result is FLAMES (legality, #1 of 13); its weakest is SALAD-Bench (mcq_representation_toxicity, #30 of 33).

Freedom. InternLM Chat 7B has an estimated Freedom rank of #307; its 90% source-sensitivity interval is #74–#563. Published Freedom evidence spans 4 eval lineages and 1 of 1 components. Its strongest relative result is SuperCLUE Safety (traditional_safety, #2 of 31); its weakest is FLAMES (legality, #13 of 13).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Fake Alignment (FINE)multiple_choice_safe_decision_rate#6 / 14↑57.33Source ↗official
Fake Alignment (FINE)open_ended_safe_response_rate#11 / 14↑92Source ↗official
FLAMESdata_protection#2 / 13↑61.84Source ↗official
FLAMESfairness#2 / 13↑44.58Source ↗official
FLAMESlegality#1 / 13↑76.09Source ↗official
FLAMESmorality#3 / 13↑51.24Source ↗official
FLAMESsafety#5 / 13↑35.9Source ↗official
SafetyBenchEM#6 / 21↑75.4Source ↗official
SafetyBenchIA#8 / 21↑79.5Source ↗official
SafetyBenchMH#6 / 21↑84.3Source ↗official
SafetyBenchOFF#12 / 21↑67.2Source ↗official
SafetyBenchPH#7 / 21↑74.15Source ↗official
SafetyBenchPP#7 / 21↑78.7Source ↗official
SafetyBenchUB#8 / 21↑64.75Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#11 / 33↑21.12Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#16 / 33↑16.61Source ↗official
SALAD-Benchattack_enhanced_malicious_use#12 / 33↑22.51Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#11 / 33↑19.9Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#9 / 33↑23.71Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#11 / 33↑22.08Source ↗official
SALAD-Benchbase_human_autonomy_integrity#10 / 33↑96.85Source ↗official
SALAD-Benchbase_information_safety_harms#15 / 33↑95.13Source ↗official
SALAD-Benchbase_malicious_use#12 / 33↑96.28Source ↗official
SALAD-Benchbase_misinformation_harms#5 / 33↑97.05Source ↗official
SALAD-Benchbase_representation_toxicity#7 / 33↑94.37Source ↗official
SALAD-Benchbase_socioeconomic_harms#15 / 33↑90.95Source ↗official
SALAD-Benchmcq_human_autonomy_integrity#30 / 33↑0Source ↗official
SALAD-Benchmcq_information_safety_harms#30 / 33↑0Source ↗official
SALAD-Benchmcq_malicious_use#30 / 33↑0.0641Source ↗official
SALAD-Benchmcq_misinformation_harms#30 / 33↑0Source ↗official
SALAD-Benchmcq_representation_toxicity#30 / 33↑0.1042Source ↗official
SALAD-Benchmcq_socioeconomic_harms#30 / 33↑0Source ↗official
SuperCLUE Safetyinstruction_attack#20 / 31↑58.62Source ↗official
SuperCLUE Safetyresponsible_ai#28 / 31↑45.45Source ↗official
SuperCLUE Safetytraditional_safety#28 / 31↑65.96Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Fake Alignment (FINE)open_ended_safe_response_rate#4 / 14↓92Source ↗official
FLAMESdata_protection#12 / 13↓61.84Source ↗official
FLAMESfairness#12 / 13↓44.58Source ↗official
FLAMESlegality#13 / 13↓76.09Source ↗official
FLAMESmorality#11 / 13↓51.24Source ↗official
FLAMESsafety#9 / 13↓35.9Source ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#22 / 33↓21.12Source ↗official
SALAD-Benchattack_enhanced_information_safety_harms#18 / 33↓16.61Source ↗official
SALAD-Benchattack_enhanced_malicious_use#22 / 33↓22.51Source ↗official
SALAD-Benchattack_enhanced_misinformation_harms#23 / 33↓19.9Source ↗official
SALAD-Benchattack_enhanced_representation_toxicity#25 / 33↓23.71Source ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#23 / 33↓22.08Source ↗official
SALAD-Benchbase_human_autonomy_integrity#24 / 33↓96.85Source ↗official
SALAD-Benchbase_information_safety_harms#19 / 33↓95.13Source ↗official
SALAD-Benchbase_malicious_use#22 / 33↓96.28Source ↗official
SALAD-Benchbase_misinformation_harms#29 / 33↓97.05Source ↗official
SALAD-Benchbase_representation_toxicity#27 / 33↓94.37Source ↗official
SALAD-Benchbase_socioeconomic_harms#19 / 33↓90.95Source ↗official
SuperCLUE Safetyinstruction_attack#9 / 31↓58.62Source ↗official
SuperCLUE Safetytraditional_safety#2 / 31↓65.96Source ↗official