← Models

Model profile

Gemini 1.5 Flash

Googledeveloper
2024-05-14release date
#113 / 346Safety rank
#450 / 662Freedom rank

Evidence summary

Safety. Gemini 1.5 Flash has an estimated Safety rank of #113; its 90% source-sensitivity interval is #49–#211. Its behavior-only rank is #123; company governance moves the combined estimate to #113. Published Safety evidence spans 17 eval lineages and 7 of 7 components. Its strongest relative result is HELM Safety (anthropic_red_team, #2 of 80); its weakest is Enkrypt AI Safety Leaderboard (toxicity_attack_non_success_rate, #226 of 268).

Freedom. Gemini 1.5 Flash has an estimated Freedom rank of #450; its 90% source-sensitivity interval is #258–#574. Published Freedom evidence spans 12 eval lineages and 1 of 1 components. Its strongest relative result is Enkrypt AI Safety Leaderboard (toxicity_attack_non_success_rate, #43 of 268); its weakest is HELM Safety (anthropic_red_team, #75 of 80).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Adversarial Robustnessscore#4 / 8↓14Source ↗official
Agent-SafetyBenchcompromise_availability#9 / 16↑30Source ↗official
Agent-SafetyBenchharmful_vulnerable_code#4 / 16↑48.4Source ↗official
Agent-SafetyBenchleak_sensitive_information#5 / 16↑39.2Source ↗official
Agent-SafetyBenchphysical_harm#6 / 16↑38.8Source ↗official
Agent-SafetyBenchproduce_unsafe_information#11 / 16↑82.4Source ↗official
Agent-SafetyBenchproperty_loss#7 / 16↑41.6Source ↗official
Agent-SafetyBenchspread_unsafe_information#4 / 16↑20.8Source ↗official
Agent-SafetyBenchviolate_law_ethics#6 / 16↑32Source ↗official
AgentDojotargeted_attack_success_rate#4 / 15↓0.0787Source ↗official
AgentDojoutility_under_attack#11 / 15↑0.333Source ↗official
AIRBench 2024 Safety Scenariossafety_scenarios#31 / 80↑0.7325Source ↗official
AnimalHarmBenchscore#3 / 10↑0.05Source ↗official
COMPL-AI AI-Identity Disclosurescore#1 / 14↑1Source ↗official
COMPL-AI LLM RuLES Multi-Turn Rule Followingscore#5 / 14↑0.5169Source ↗official
COMPL-AI TensorTrust Goal-Hijacking Resistancescore#6 / 13↑0.4923Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#210 / 270↑12.14Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#147 / 270↑86.67Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#184 / 270↑55.56Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#226 / 268↑90.59Source ↗official
FinEval Financial Security Knowledgefinancial_security_accuracy_pct#6 / 19↑77.5Source ↗official
FORTRESSaverage_risk_score#46 / 60↓50.61Source ↗official
FORTRESSover_refusal_score#27 / 59↓4.45Source ↗official
HELM Safetyanthropic_red_team#2 / 80↑0.999Source ↗official
HELM Safetybbq#32 / 80↑0.947Source ↗official
HELM Safetyharmbench#35 / 80↑0.8Source ↗official
HELM Safetysimple_safety_tests#58 / 80↑0.97Source ↗official
HELM Safetyxstest#66 / 80↑0.921Source ↗official
Large-scale Moral Machine experiment on LLMshuman_choice_distance#23 / 39↓1.116Source ↗official
OR-Benchover_refusal_rate#17 / 25↓84.3Source ↗official
OR-Benchtoxic_acceptance_rate#8 / 25↓1.2Source ↗official
SORRY-Benchavg#5 / 51↓0.08Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Adversarial Robustnessscore#5 / 8↑14Source ↗official
Agent-SafetyBenchproduce_unsafe_information#6 / 16↓82.4Source ↗official
AIRBench 2024 Safety Scenariossafety_scenarios#50 / 80↓0.7325Source ↗official
COMPL-AI LLM RuLES Multi-Turn Rule Followingscore#10 / 14↓0.5169Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#122 / 270↓86.67Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#82 / 270↓55.56Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#43 / 268↓90.59Source ↗official
FORTRESSaverage_risk_score#15 / 60↑50.61Source ↗official
FORTRESSover_refusal_score#27 / 59↓4.45Source ↗official
HELM Safetyanthropic_red_team#75 / 80↓0.999Source ↗official
HELM Safetyharmbench#46 / 80↓0.8Source ↗official
HELM Safetysimple_safety_tests#20 / 80↓0.97Source ↗official
HELM Safetyxstest#66 / 80↑0.921Source ↗official
OR-Benchover_refusal_rate#17 / 25↓84.3Source ↗official
OR-Benchtoxic_acceptance_rate#18 / 25↑1.2Source ↗official
SORRY-Benchavg#45 / 51↑0.08Source ↗official