← Models

Model profile

Gemma 7B It

Googledeveloper
2024-02-21release date
#167 / 267overall rank
7eval lineages

Evidence summary

Gemma 7B It has an estimated overall rank of #167; its 90% source-sensitivity interval is #48–#208. Its behavior-only rank is #173; company governance moves the combined estimate to #167. Published evidence spans 7 evals and 7 of 7 behavior components. Its strongest relative result is SALAD-Bench (base_representation_toxicity, #6 of 33); its weakest is Enkrypt AI Safety Leaderboard (toxicity_attack_non_success_rate, #242 of 258).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#88 / 26020.41↑ higherSource ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#225 / 26078.83↑ higherSource ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#181 / 26054.44↑ higherSource ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#242 / 25877.05↑ higherSource ↗official
Large-scale Moral Machine experiment on LLMshuman_choice_distance#32 / 391.314↓ lowerSource ↗official
Microsoft Phi Safety Panelsharmful_continuation#6 / 100.013↓ lowerSource ↗official
Microsoft Phi Safety Panelsharmful_summarization#3 / 100.103↓ lowerSource ↗official
Microsoft Phi Safety Panelsjailbreak#4 / 100.114↓ lowerSource ↗official
Microsoft Phi Safety Panelsthird_party_harm#8 / 100.383↓ lowerSource ↗official
OR-Benchover_refusal_rate#7 / 2526.3↓ lowerSource ↗official
OR-Benchtoxic_acceptance_rate#17 / 2514.5↓ lowerSource ↗official
S-Evalbase_en_overall#10 / 2261.8↑ higherSource ↗official
SALAD-Benchattack_enhanced_human_autonomy_integrity#16 / 3313.36↑ higherSource ↗official
SALAD-Benchattack_enhanced_information_safety_harms#11 / 3322.8↑ higherSource ↗official
SALAD-Benchattack_enhanced_malicious_use#19 / 339.95↑ higherSource ↗official
SALAD-Benchattack_enhanced_misinformation_harms#12 / 3318.91↑ higherSource ↗official
SALAD-Benchattack_enhanced_representation_toxicity#15 / 3317.56↑ higherSource ↗official
SALAD-Benchattack_enhanced_socioeconomic_harms#18 / 3312.12↑ higherSource ↗official
SALAD-Benchbase_human_autonomy_integrity#17 / 3394.82↑ higherSource ↗official
SALAD-Benchbase_information_safety_harms#9 / 3397.49↑ higherSource ↗official
SALAD-Benchbase_malicious_use#17 / 3393.54↑ higherSource ↗official
SALAD-Benchbase_misinformation_harms#12 / 3395.57↑ higherSource ↗official
SALAD-Benchbase_representation_toxicity#6 / 3394.42↑ higherSource ↗official
SALAD-Benchbase_socioeconomic_harms#22 / 3386.13↑ higherSource ↗official
SALAD-Benchmcq_human_autonomy_integrity#15 / 3340.56↑ higherSource ↗official
SALAD-Benchmcq_information_safety_harms#12 / 3340.56↑ higherSource ↗official
SALAD-Benchmcq_malicious_use#14 / 3340.38↑ higherSource ↗official
SALAD-Benchmcq_misinformation_harms#16 / 3338.1↑ higherSource ↗official
SALAD-Benchmcq_representation_toxicity#16 / 3338.85↑ higherSource ↗official
SALAD-Benchmcq_socioeconomic_harms#14 / 3338.89↑ higherSource ↗official
SORRY-Benchavg#16 / 510.18↓ lowerSource ↗official