← Evals

Evaluation profile

BioSecBench-Refusal V2

2sub-evals
0.177%Safety weight
0.636%Freedom weight
2components

Weights below are portfolio-specific global index weights.

Model score (higher is better)Predicted score

About this eval

Deployed-agent refusal of concealed biological hazards and non-overrefusal of legitimate biological work.

Included in the behavior ranking.

Sub-evals

MeasureComponentDirection by rankingSafety weightFreedom weight
redteam_refusal_pctbenchmarks-bio-security/benchmarks-bio-security.csv:redteam_refusal_pctMeasures whether the deployed agent refuses requests concealing biological hazards.Safety: ordinary_harm_misuse_resistance:1.000 · Freedom: user_direction_following:1.000biosecbench-refusalSafety: higher · Freedom: lower0.0223%0.318%
routine_compliance_pctbenchmarks-bio-security/benchmarks-bio-security.csv:routine_compliance_pctMeasures whether the deployed agent allows legitimate biological research requests without over-refusing them.Safety: benign_helpfulness_non_overrefusal:1.000 · Freedom: user_direction_following:1.000biosecbench-refusalSafety: higher · Freedom: higher0.155%0.318%

redteam_refusal_pct

Measures whether the deployed agent refuses requests concealing biological hazards.

RankModelValueRelative performanceProvenance
1gpt-6-astra96.01official
2claude-opus-591.67official
3grok-4.779.71official
4gemini-3.8-flash77.54official
5gemini-3.7-flash63.64official
6grok-4.642.39official

routine_compliance_pct

Measures whether the deployed agent allows legitimate biological research requests without over-refusing them.

RankModelValueRelative performanceProvenance
1grok-4.753.59official
2grok-4.653.57official
3gemini-3.7-flash49.72official
4gemini-3.8-flash36.61official
5gpt-6-astra22.4official
6claude-opus-518.58official