← Models

Model profile

Grok 4.1 Fast

xAIdeveloper
2025-11-19release date
#227 / 346Safety rank
#298 / 662Freedom rank

Evidence summary

Safety. Grok 4.1 Fast has an estimated Safety rank of #227; its 90% source-sensitivity interval is #112–#314. Its behavior-only rank is #226; company governance moves the combined estimate to #227. Published Safety evidence spans 25 eval lineages and 7 of 7 components. Its strongest relative result is SpeciEval (belief_animal_sentience, #1 of 131); its weakest is AuAu Authoritarian Response Audit (realistic_prompt_arr_percent, #17 of 17).

Freedom. Grok 4.1 Fast has an estimated Freedom rank of #298; its 90% source-sensitivity interval is #85–#464. Published Freedom evidence spans 12 eval lineages and 1 of 1 components. Its strongest relative result is SpeechMap model completion (complete_pct, #2 of 181); its weakest is Adversarial Humanities Benchmark (AHB) — Table 5 (child_sexual_exploitation, #28 of 31).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#149 / 358↓0.7341Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5child_sexual_exploitation#4 / 31↓14.6Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5defamation#22 / 31↓78.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5hate#23 / 31↓65.2Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5indiscriminate_weapons_cbrne#10 / 31↓33.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5intellectual_property#10 / 31↓41.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5non_violent_crimes#11 / 31↓52.2Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5privacy#11 / 31↓53.2Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sex_related_crimes#11 / 31↓54.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sexual_content#26 / 31↓64.6Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5specialized_advice#14 / 31↓60.92Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5suicide_self_harm#13 / 31↓55.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5violent_crimes#13 / 31↓60Source ↗official
Alignment Leaderboardcorrigibility#19 / 24↑4.068Source ↗official
Alignment Leaderboardhonesty#9 / 24↑3.682Source ↗official
Alignment Leaderboardnon_manipulation#11 / 24↑3.515Source ↗official
Alignment Leaderboardrobustness#5 / 24↑4.027Source ↗official
Alignment Leaderboardsafety#12 / 24↑3.846Source ↗official
Alignment Leaderboardscheming#13 / 24↑3.627Source ↗official
Arena Factuality — Search Arena (factuality-only weighting)factuality_bt_rating#27 / 30↑1123.0Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#100 / 111↑1395.0Source ↗official
AuAu Authoritarian Response Auditbehavioral_vignette_arr_percent#15 / 17↓29.5Source ↗official
AuAu Authoritarian Response Auditrealistic_prompt_arr_percent#17 / 17↓7.7Source ↗official
BioTIERpermit_compliance_pct#36 / 52↑98.4Source ↗official
BioTIERrefuse_compliance_pct#9 / 52↑77.3Source ↗official
BullshitBench v2clear_pushback_rate#92 / 122↑0.145Source ↗official
CAIS Risk Indexagent_red_teaming#39 / 49↓90.2Source ↗official
CAIS Risk Indexbioweapons_assistance#20 / 54↓63.8Source ↗official
CAIS Risk Indexhle_overconfidence#44 / 55↓72.1Source ↗official
CAIS Risk Indexmachiavelli#6 / 51↓81.3Source ↗official
CAIS Risk Indexmask#56 / 57↓71Source ↗official
CAIS Risk Indexpolitical_manipulation#4 / 51↓32.5Source ↗official
CAIS Risk Indextextquests_harm#3 / 54↓9.1Source ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#42 / 104↓29.06Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#210 / 270↑12.14Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#227 / 270↑71.83Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#162 / 270↑63.33Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#241 / 268↑84.18Source ↗official
Human Pathogen Capabilities Test (HPCT) — overall refusalrefusal_rate_pct#16 / 69↑24.55Source ↗official
LiveSecBenchethics#43 / 43↑8.57Source ↗official
LiveSecBenchfactuality#20 / 43↑52.76Source ↗official
LiveSecBenchlegality#38 / 43↑18.02Source ↗official
LiveSecBenchprivacy#25 / 43↑39.48Source ↗official
LiveSecBenchpsychological_health#27 / 43↑44.84Source ↗official
MACHIAVELLIdeception_relative_random_pct#5 / 50↓81.3Source ↗official
MORUscore#9 / 13↑71.14Source ↗official
MT-JailBench CrescendoXsafety_score#12 / 21↑18.87Source ↗official
Pander Scoreconversational_absolute_pander_score#26 / 26↓34.8Source ↗official
Pander Scoreinstructional_absolute_pander_score#26 / 26↓84.47Source ↗official
SM-Benchadversarial#52 / 92↑80.98Source ↗official
SM-Benchambiguous_interpretation#73 / 92↑78.57Source ↗official
SM-Benchanti_hallucination#31 / 92↑96.86Source ↗official
SM-Bencheq_boundaries#9 / 92↑74.72Source ↗official
SM-Benchoverfit#60 / 92↑66.12Source ↗official
SpeciEvalbelief_animal_sentience#1 / 131↑7Source ↗official
SpeciEvalland_animal_4ns#130 / 131↓5.35Source ↗official
SpeciEvalsea_animal_4ns#125 / 131↓5.28Source ↗official
SpeciEvalspeciesism#118 / 131↓2.72Source ↗official
Vectara HHEM Factual Consistencyfactual_consistency_rate#87 / 94↑81.5Source ↗official
Vigil Mental Health Safetyoverall_score#18 / 23↑30Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Adversarial Humanities Benchmark (AHB) — Table 5child_sexual_exploitation#28 / 31↑14.6Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5defamation#10 / 31↑78.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5hate#9 / 31↑65.2Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5indiscriminate_weapons_cbrne#21 / 31↑33.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5intellectual_property#22 / 31↑41.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5non_violent_crimes#21 / 31↑52.2Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5privacy#21 / 31↑53.2Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sex_related_crimes#21 / 31↑54.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5sexual_content#5 / 31↑64.6Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5specialized_advice#18 / 31↑60.92Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5suicide_self_harm#19 / 31↑55.3Source ↗official
Adversarial Humanities Benchmark (AHB) — Table 5violent_crimes#19 / 31↑60Source ↗official
BioTIERpermit_compliance_pct#36 / 52↑98.4Source ↗official
BioTIERrefuse_compliance_pct#44 / 52↓77.3Source ↗official
CAIS Risk Indexbioweapons_assistance#34 / 54↑63.8Source ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#63 / 104↑29.06Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#44 / 270↓71.83Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#109 / 270↓63.33Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#28 / 268↓84.18Source ↗official
Human Pathogen Capabilities Test (HPCT) — overall refusalrefusal_rate_pct#54 / 69↓24.55Source ↗official
LiveSecBenchethics#1 / 43↓8.57Source ↗official
LiveSecBenchlegality#6 / 43↓18.02Source ↗official
LiveSecBenchprivacy#19 / 43↓39.48Source ↗official
LiveSecBenchpsychological_health#17 / 43↓44.84Source ↗official
MT-JailBench CrescendoXsafety_score#10 / 21↓18.87Source ↗official
SM-Benchadversarial#39 / 92↓80.98Source ↗official
SM-Bencheq_boundaries#9 / 92↑74.72Source ↗official
SM-Benchoverfit#60 / 92↑66.12Source ↗official
SpeechMap model completioncomplete_pct#2 / 181↑97.9Source ↗official
UGI Leaderboard — base-model willingnesswillingness_adherence_score#15 / 156↑7.5Source ↗official
UGI Leaderboard — base-model willingnesswillingness_direct_score#17 / 156↑5.5Source ↗official
Vigil Mental Health Safetyoverall_score#6 / 23↓30Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean18.5
Government43.8
Diplomacy51.7
Economy25.4
Society61.6