Evaluation profile
GPT-5.6 system card — first-person fairness
1sub-evals
0.0563%total index weight
1components
Within-component eval weight: Fairness 0.563%.
Model score (lower is better)Predicted score
About this eval
Harmful stereotyping differences in responses conditioned on names statistically associated with male versus female users.
Included in the behavior ranking.
Sub-evals
| Measure | Component | Direction | Total index weight | Within-component weight |
|---|---|---|---|---|
| harm_overall_pctgpt56-system-card/first-person-fairness.csv:harm_overall_pctMeasures whether otherwise equivalent first-person prompts receive different answers when the associated name is male rather than female. | fairness_nondiscrimination:1.000gpt56-system-card | Lower is better | 0.0563% | Fairness 0.563% |
harm_overall_pct
Measures whether otherwise equivalent first-person prompts receive different answers when the associated name is male rather than female.
| Rank | Model | Value | Relative performance | Provenance |
|---|---|---|---|---|
| 1 | gpt-5.6-luna | 0.61 | official | |
| 2 | gpt-5.4 | 0.88 | official | |
| 2 | gpt-5.6-terra | 0.88 | official | |
| 4 | gpt-5.6-sol | 0.98 | official | |
| 5 | gpt-5.2 | 1 | official | |
| 6 | gpt-5.5 | 1.12 | official | |
| 7 | gpt-5.1 | 1.28 | official |