← Evals

Evaluation profile

Pander Score

2sub-evals
0.506%Safety weight
0%Freedom weight
1components

Weights below are portfolio-specific global index weights.

Model score (lower is better)Predicted score

About this eval

Magnitude of epistemically poor response-belief movement with user belief, whether deferential (pandering) or oppositional (contrarian).

Included in the behavior ranking.

Sub-evals

MeasureComponentDirection by rankingSafety weightFreedom weight
conversational_absolute_pander_scorepander-score/pander-score.csv:conversational_absolute_pander_scoreMeasures how strongly a model's expressed belief moves either with or against the user's stated belief in conversational questions.Safety: truthfulness_honesty:1.000pander-scoreSafety: lower0.253%—
instructional_absolute_pander_scorepander-score/pander-score.csv:instructional_absolute_pander_scoreMeasures how strongly a model's expressed belief moves either with or against assumptions embedded in user instructions.Safety: truthfulness_honesty:1.000pander-scoreSafety: lower0.253%—

conversational_absolute_pander_score

Measures how strongly a model's expressed belief moves either with or against the user's stated belief in conversational questions.

RankModelValueRelative performanceProvenance
1claude-opus-51.028official
2claude-fable-51.093official
3claude-fable-5.11.355official
4gpt-6-astra1.557official
5claude-opus-4.63.487official
6claude-sonnet-4.63.936official
7muse-spark-1.14.584official
8gpt-5.6-sol6.553official
9kimi-k37.048official
10claude-sonnet-57.263official
11gpt-5.6-terra9.7official
12glm-5.310.37official
13grok-4.613.82official
14gpt-5.414.31official
15gemini-3.8-flash15.3official
16gemini-3.7-flash15.68official
17grok-4.516.28official
18gpt-5.4-mini16.43official
19inkling18.37official
20gemini-3.6-flash18.48official
21grok-4.2019.26official
22gemini-3.1-pro-preview22.66official
23gemini-3.5-flash23.33official
24gemini-3-flash-preview25.97official
25glm-5.228.23official
26grok-4.1-fast34.8official

instructional_absolute_pander_score

Measures how strongly a model's expressed belief moves either with or against assumptions embedded in user instructions.

RankModelValueRelative performanceProvenance
1gpt-6-astra2.589official
2claude-opus-55.172official
3claude-sonnet-4.611.79official
4claude-fable-5.113.13official
5claude-opus-4.616.83official
6muse-spark-1.116.95official
7gpt-5.6-sol17.75official
8claude-fable-518.59official
9kimi-k323.5official
10gpt-5.6-terra25.68official
11glm-5.327.24official
12claude-sonnet-529.5official
13gpt-5.432.4official
14grok-4.633.62official
15gpt-5.4-mini38.82official
16grok-4.541.27official
17inkling45.4official
18gemini-3.7-flash65.05official
19gemini-3.6-flash66.01official
20gemini-3.8-flash66.04official
21grok-4.2068.73official
22gemini-3.5-flash69.48official
23glm-5.269.88official
24gemini-3-flash-preview70.12official
25gemini-3.1-pro-preview70.61official
26grok-4.1-fast84.47official