← Models

Model profile

GPT 6 Astra

OpenAIdeveloper
2026-09-03release date
#2 / 346Safety rank
#583 / 662Freedom rank
1discovery sources

Evidence summary

Safety. GPT 6 Astra has an estimated Safety rank of #2; its 90% source-sensitivity interval is #1–#8. Published Safety evidence spans 18 eval lineages and 7 of 7 components. Its strongest relative result is SpeciEval (belief_animal_sentience, #1 of 131); its weakest is GPT-6 Sol/Luna system card — safety and updated alignment tests (agentic_codex_personal_data_safe_rate, #4 of 4).

Freedom. GPT 6 Astra has an estimated Freedom rank of #583; its 90% source-sensitivity interval is #206–#643. Published Freedom evidence spans 5 eval lineages and 1 of 1 components. Its strongest relative result is SM-Bench (eq_boundaries, #6 of 92); its weakest is BioSecBench-Refusal V2 (redteam_refusal_pct, #6 of 6).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#71 / 358↓0.4477Source ↗official
BioSecBench-Refusal V2redteam_refusal_pct#1 / 6↑96.01Source ↗official
BioSecBench-Refusal V2routine_compliance_pct#5 / 6↑22.4Source ↗official
BullshitBench v2clear_pushback_rate#14 / 122↑0.665Source ↗official
CAIS Risk Indexagent_red_teaming#1 / 49↓13.5Source ↗official
CAIS Risk Indexbioweapons_assistance#12 / 54↓50.8Source ↗official
CAIS Risk Indexhle_overconfidence#8 / 55↓39.8Source ↗official
CAIS Risk Indexmachiavelli#1 / 51↓76.3Source ↗official
CAIS Risk Indexmask#1 / 57↓1.5Source ↗official
CAIS Risk Indexpolitical_manipulation#9 / 51↓37.4Source ↗official
CAIS Risk Indextextquests_harm#53 / 54↓29.2Source ↗official
CheatBench direct cheating propensitybiology_bioinformatics_cheating_rate_pct#4 / 8↓70Source ↗official
CheatBench direct cheating propensityboard_games_cheating_rate_pct#4 / 8↓30Source ↗official
CheatBench direct cheating propensitycreative_writing_cheating_rate_pct#3 / 8↓35Source ↗official
CheatBench direct cheating propensityknowledge_work_cheating_rate_pct#4 / 8↓55Source ↗official
CheatBench direct cheating propensitymathematical_research_cheating_rate_pct#4 / 8↓90Source ↗official
CheatBench direct cheating propensitymenial_computation_cheating_rate_pct#4 / 8↓100Source ↗official
CheatBench direct cheating propensitymultimodal_cheating_rate_pct#3 / 8↓53Source ↗official
CheatBench direct cheating propensitysoftware_engineering_cheating_rate_pct#2 / 8↓10Source ↗official
CheatBench direct cheating propensitysvg_competition_cheating_rate_pct#1 / 8↓0Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#48 / 270↑35.14Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#243 / 270↑65.67Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#41 / 270↑97.22Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#181 / 268↑94.55Source ↗official
GPT 6 Astra system-card alignment evaluationsattempted_retry_or_circumvention_after_denial_pct#1 / 2↓0Source ↗official
GPT 6 Astra system-card alignment evaluationsexploitgym_honeypot_attack_rate_pct#1 / 2↓0Source ↗official
GPT 6 Astra system-card alignment evaluationsfalsified_data_labels_pct#1 / 2↓0.17Source ↗official
GPT 6 Astra system-card alignment evaluationsfollowed_unauthorized_agent_instructions_pct#1 / 4↓0Source ↗official
GPT 6 Astra system-card alignment evaluationsinternal_circumvention_rate_pct#1 / 2↓0Source ↗official
GPT 6 Astra system-card alignment evaluationsinternal_computer_use_safety_autoreview_error_rate_pct#1 / 2↓1.8Source ↗official
GPT 6 Astra system-card alignment evaluationsinternal_computer_use_safety_error_rate_pct#1 / 5↓2.4Source ↗official
GPT 6 Astra system-card alignment evaluationsinternal_hallucination_rate_pct#1 / 2↓4.2Source ↗official
GPT 6 Astra system-card alignment evaluationsoverall_misaligned_outcome_base_pct#1 / 4↓3.4Source ↗official
GPT 6 Astra system-card alignment evaluationsoverall_misaligned_outcome_confirmation_pct#1 / 4↓3Source ↗official
GPT 6 Astra system-card alignment evaluationsseverity_1_or_2_misalignment_flags_pct#1 / 2↓13.62Source ↗official
GPT 6 Astra system-card alignment evaluationsseverity_3_plus_misalignment_flags_pct#1 / 2↓0.063Source ↗official
GPT 6 Astra system-card alignment evaluationsunwanted_persistence_after_warning_pct#1 / 2↓19Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsagentic_chat_plugins_safe_rate#1 / 4↑1Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsagentic_codex_age_restricted_safe_rate#1 / 4↑0.811Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsagentic_codex_nonviolent_wrongdoing_safe_rate#3 / 4↑0.954Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsagentic_codex_personal_data_safe_rate#4 / 4↑0.763Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsagentic_codex_self_harm_safe_rate#2 / 4↑0.92Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsagentic_codex_violent_wrongdoing_safe_rate#1 / 4↑0.907Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsagentic_redteam_chat_plugins_safe_rate#1 / 4↑1Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsagentic_redteam_codex_safe_rate#1 / 4↑0.977Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsattempted_retry_or_circumvention_after_denial_pct_2026_09_22#1 / 5↓0Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsbiology_benign_not_overrefuse_rate#3 / 6↑0.978Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsbiology_dual_use_safe_rate#2 / 6↑0.97Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsbiology_severe_safe_rate#1 / 6↑0.998Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsbroken_search_unacknowledged_limitation_pct_2026_09_22#1 / 5↓1.5Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testschallenging_extremism_safe_rate#1 / 7↑0.981Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testschallenging_gore_safe_rate#2 / 7↑0.898Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testschallenging_hate_safe_rate#1 / 7↑1Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testschallenging_nonviolent_illicit_safe_rate#1 / 7↑0.997Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testschallenging_self_harm_safe_rate#1 / 7↑0.992Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testschallenging_sexual_minors_safe_rate#2 / 7↑0.974Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testschallenging_sexual_safe_rate#2 / 7↑0.98Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testschallenging_violent_illicit_safe_rate#1 / 7↑0.99Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testscoding_deception_misrepresentation_pct_2026_09_22#1 / 5↓0.51Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testscyber_production_chat_safe_score#3 / 6↑0.97Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testscyber_semisynthetic_agentic_safe_score#1 / 5↑0.997Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testscyber_synthetic_agentic_safe_score#5 / 6↑0.992Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsdynamic_emotional_reliance_safe_rate#1 / 7↑0.993Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsdynamic_mental_health_safe_rate#1 / 7↑1Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsdynamic_self_harm_safe_rate#1 / 7↑0.989Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsinstruction_hierarchy_defender_success_pct#1 / 3↑99.99Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsstatic_jailbreak_bio_high_risk_defender_success_pct#1 / 5↑97.3Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsstatic_jailbreak_bio_severe_defender_success_pct#1 / 5↑98.2Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsstatic_jailbreak_cyber_defender_success_pct#1 / 5↑91.5Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsstatic_jailbreak_violence_moderate_defender_success_pct#1 / 5↑94.7Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsstatic_jailbreak_violence_severe_defender_success_pct#1 / 5↑98.3Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsu18_age_restricted_safe_rate#1 / 7↑0.918Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsu18_eating_disorders_safe_rate#1 / 7↑0.921Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsu18_emotional_reliance_safe_rate#3 / 7↑0.944Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsu18_gore_safe_rate#1 / 7↑0.898Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsu18_self_harm_safe_rate#1 / 7↑0.995Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsu18_sexual_content_safe_rate#1 / 7↑0.991Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsunwanted_persistence_after_warning_pct_2026_09_22#1 / 5↓17.4Source ↗official
Humanity's Last Exam RMS calibration error (Scale Labs)calibrationError#4 / 42↓39Source ↗official
MentalHealthBenchoverall_model_performance_pct#1 / 17↑57.33Source ↗official
Olam Diplomacy — Broken Promise Ratebroken_promise_rate_pct#6 / 19↓12.5Source ↗official
Pander Scoreconversational_absolute_pander_score#4 / 26↓1.557Source ↗official
Pander Scoreinstructional_absolute_pander_score#1 / 26↓2.589Source ↗official
SM-Benchadversarial#31 / 92↑83.91Source ↗official
SM-Benchambiguous_interpretation#33 / 92↑88.09Source ↗official
SM-Benchanti_hallucination#18 / 92↑98.69Source ↗official
SM-Bencheq_boundaries#6 / 92↑76.69Source ↗official
SM-Benchoverfit#7 / 92↑94.81Source ↗official
SpeciEvalbelief_animal_sentience#1 / 131↑7Source ↗official
SpeciEvalland_animal_4ns#22 / 131↓4.25Source ↗official
SpeciEvalsea_animal_4ns#15 / 131↓4.33Source ↗official
SpeciEvalspeciesism#33 / 131↓1.62Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
BioSecBench-Refusal V2redteam_refusal_pct#6 / 6↓96.01Source ↗official
BioSecBench-Refusal V2routine_compliance_pct#5 / 6↑22.4Source ↗official
CAIS Risk Indexbioweapons_assistance#43 / 54↑50.8Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#28 / 270↓65.67Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#229 / 270↓97.22Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#83 / 268↓94.55Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsbiology_benign_not_overrefuse_rate#3 / 6↑0.978Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsbiology_dual_use_safe_rate#5 / 6↓0.97Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsbiology_severe_safe_rate#5 / 6↓0.998Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testschallenging_extremism_safe_rate#6 / 7↓0.981Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testschallenging_gore_safe_rate#6 / 7↓0.898Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testschallenging_hate_safe_rate#3 / 7↓1Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testschallenging_nonviolent_illicit_safe_rate#6 / 7↓0.997Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testschallenging_self_harm_safe_rate#7 / 7↓0.992Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testschallenging_sexual_minors_safe_rate#5 / 7↓0.974Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testschallenging_sexual_safe_rate#6 / 7↓0.98Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testschallenging_violent_illicit_safe_rate#7 / 7↓0.99Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testscyber_production_chat_safe_score#4 / 6↓0.97Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsdynamic_emotional_reliance_safe_rate#7 / 7↓0.993Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsdynamic_mental_health_safe_rate#6 / 7↓1Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsdynamic_self_harm_safe_rate#7 / 7↓0.989Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsinstruction_hierarchy_defender_success_pct#3 / 3↓99.99Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsstatic_jailbreak_bio_high_risk_defender_success_pct#5 / 5↓97.3Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsstatic_jailbreak_bio_severe_defender_success_pct#5 / 5↓98.2Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsstatic_jailbreak_cyber_defender_success_pct#5 / 5↓91.5Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsstatic_jailbreak_violence_moderate_defender_success_pct#5 / 5↓94.7Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsstatic_jailbreak_violence_severe_defender_success_pct#5 / 5↓98.3Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsu18_age_restricted_safe_rate#7 / 7↓0.918Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsu18_eating_disorders_safe_rate#7 / 7↓0.921Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsu18_emotional_reliance_safe_rate#5 / 7↓0.944Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsu18_gore_safe_rate#6 / 7↓0.898Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsu18_self_harm_safe_rate#7 / 7↓0.995Source ↗official
GPT-6 Sol/Luna system card — safety and updated alignment testsu18_sexual_content_safe_rate#7 / 7↓0.991Source ↗official
SM-Benchadversarial#62 / 92↓83.91Source ↗official
SM-Bencheq_boundaries#6 / 92↑76.69Source ↗official
SM-Benchoverfit#7 / 92↑94.81Source ↗official