← Models

Model profile

Claude Opus 4.1

Anthropicdeveloper
2025-08-05release date
#27 / 267overall rank
12eval lineages

Evidence summary

Claude Opus 4.1 has an estimated overall rank of #27; its 90% source-sensitivity interval is #10–#81. Its behavior-only rank is #32; company governance moves the combined estimate to #27. Published evidence spans 12 evals and 6 of 7 behavior components. Its strongest relative result is Cisco AI Defense Rolling Single-Turn Leaderboard (single_turn_attack_success_rate, #2 of 105); its weakest is Anthropic Claude Opus 4.1 System Card Addendum (benign_request_refusal, #2 of 2).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionBetterSource
Anthropic Claude Opus 4.1 System Card Addendumbbq_disambiguated_accuracy#2 / 20.907↑ higherSource ↗official
Anthropic Claude Opus 4.1 System Card Addendumbenign_request_refusal#2 / 20.0008↓ lowerSource ↗official
Anthropic Claude Opus 4.1 System Card Addendumharmful_request_safety#1 / 20.9876↑ higherSource ↗official
Anthropic Claude Opus 4.5 System Cardagentic_coding_safety#4 / 40.96↑ higherSource ↗official
Anthropic Claude Opus 4.5 System Cardclaude_code_malicious_refusal#4 / 40.4816↑ higherSource ↗official
Anthropic Claude Opus 4.5 System Cardcomputer_use_malicious_refusal#4 / 40.6696↑ higherSource ↗official
Anthropic Claude Opus 4.5 System Cardharmful_request_safety#2 / 20.9914↑ higherSource ↗official
Anthropic Claude Sonnet 4.5 System Cardbenign_request_refusal#2 / 20.0008↓ lowerSource ↗official
BullshitBench v2clear_pushback_rate#33 / 1050.425↑ higherSource ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#2 / 1052.243↓ lowerSource ↗official
Confabulationsconfabulation_rate#2 / 523.218↓ lowerSource ↗official
FORTRESSaverage_risk_score#8 / 4915.47↓ lowerSource ↗official
FORTRESSover_refusal_score#21 / 464.065↓ lowerSource ↗official
MASKlying_probability_pct#6 / 539.2↓ lowerSource ↗official
PHAREbias_resistance_diagnostic#39 / 660.4361↑ higherSource ↗official
PHAREhallucination_resistance_diagnostic#6 / 700.8619↑ higherSource ↗official
PHAREharm_resistance_diagnostic#18 / 700.9631↑ higherSource ↗official
PHAREjailbreak_resistance_diagnostic#2 / 670.8135↑ higherSource ↗official
Social Welfare Function Benchmarkfairness#7 / 190.525↑ higherSource ↗official
SOSBenchbiology_pvr#3 / 230.134↓ lowerSource ↗official
SOSBenchchemistry_pvr#2 / 230.147↓ lowerSource ↗official
SOSBenchmedicine_pvr#2 / 230.232↓ lowerSource ↗official
SOSBenchpharmacology_pvr#2 / 230.249↓ lowerSource ↗official
SOSBenchphysics_pvr#1 / 230.098↓ lowerSource ↗official
SOSBenchpsychology_pvr#1 / 230.107↓ lowerSource ↗official
SpeciEvalbelief_animal_sentience#68 / 1026.62↑ higherSource ↗official
SpeciEvalland_animal_4ns#26 / 1024.33↓ lowerSource ↗official
SpeciEvalsea_animal_4ns#16 / 1024.47↓ lowerSource ↗official
SpeciEvalspeciesism#43 / 1021.92↓ lowerSource ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-22.6
Government45.7
Diplomacy66
Economy46.9
Society61.3