← Models

Model profile

OLMo 7B Instruct HF

Ai2developer
2024-06-04release date
Not rankedSafety rank
#132 / 662Freedom rank

Evidence summary

Safety. OLMo 7B Instruct HF does not meet the evidence gate for a Safety rank. Published Safety evidence spans 1 eval lineages and 3 of 7 components. Its strongest relative result is Enkrypt AI Safety Leaderboard (bias_attack_non_success_rate, #124 of 270); its weakest is Enkrypt AI Safety Leaderboard (toxicity_attack_non_success_rate, #264 of 268).

Freedom. OLMo 7B Instruct HF has an estimated Freedom rank of #132; its 90% source-sensitivity interval is #2–#549. Published Freedom evidence spans 1 eval lineages and 1 of 1 components. Its strongest relative result is Enkrypt AI Safety Leaderboard (toxicity_attack_non_success_rate, #5 of 268); its weakest is Enkrypt AI Safety Leaderboard (cbrn_attack_non_success_rate, #64 of 270).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#124 / 270↑16.8Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#205 / 270↑79Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#263 / 270↑25.56Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#264 / 268↑59.55Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#64 / 270↓79Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#7 / 270↓25.56Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#5 / 268↓59.55Source ↗official