Developer
Meta
36 indexed models; 17 currently meet the evidence threshold for the overall ranking. Together they have results from 64 evaluations.
Company governance evidence
Meta is represented at -1.25 SD relative to the matched Future of Life Institute edition. Provenance: Direct. This company-level evidence contributes 10% of overall model rank.
Models by Meta
Evaluations covering Meta models (64)
AA-Omniscience · AbstentionBench · Adversarial Robustness · Agent-SafetyBench · AgentDojo · AgentHarm · AILuminate General Purpose AI Chat · AIRBench 2024 Safety Scenarios · Alignment Leaderboard · AnimalHarmBench · Anthropic Agentic Misalignment — blackmail · Anthropic Agentic Misalignment — corporate espionage · BlueBench AttaQ-100 · BullshitBench v2 · CAIS Risk Index · CASE-Bench · ChineseSafe · Cisco AI Defense Rolling Single-Turn Leaderboard · Confabulations · Contextual MoralChoice · DecodingTrust · Do-Not-Answer · DSPSafeBench · DystopiaBench · Enkrypt AI Safety Leaderboard · FORTRESS · Gray Swan indirect prompt injection (15 attempts) · HarmBench · HELM Classic RealToxicityPrompts · HELM Safety · HUMAINE Trust, Ethics and Safety · JailBench · Large-scale Moral Machine experiment on LLMs · LiveSecBench · LLM Ethics Benchmark · MACHIAVELLI · Manager Coercion Bench · MANTA · MASK · Microsoft Phi Safety Panels · ODCV-Bench · Open LLM Safety Index · OR-Bench · PandaBench JBB direct-request panel · PHARE · PropensityBench · RefusalBench · S-Eval · SafeArena · SafetyBench · SALAD-Bench · Shell · SORRY-Bench · SOSBench · SpeciesismBench · SpeciEval · SYCON Bench · TAC · TrustLLM contemporary collapsed application · TukaBench · UAVBench safety-critical decision recognition · VETO Misfired Alignment · Vigil Mental Health Safety · XSTest