AI Security Leaderboard – comparing cyber and CBRN safeguards
AI Security Leaderboard – comparing cyber and CBRN safeguards
There's no shortage of leaderboards for model capabilities - but the security of models is becoming increasingly relevant, from the risk of an AI agent processing unsanitized input being hijacked to models being pulled due to cybersecurity jailbreaks. We developed an automated test suite that runs models through 1500 automatically generated jailbreak attempts and measures the number of universal jailbreaks: prompts that elicit compliant, detailed responses to >75% clearly harmful questions within a domain (like offensive cybersecurity). We find a big gap between the most robust models -- Fable 5 and GPT-5.6 Sol -- and other leading frontier models -- Gemini 3.1 Pro and Grok 4.5. This is v1.0 and we plan to update with new attacks and broader datasets in the future; we'd love to hear from HN what would be useful in your work!
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Correct prediction on native model
Similar products
Cyber Security
Your AI Security Engineer
Your AI Security Engineer
AI Security Newsletter
AI Security Breaches, the problems and it's causes
AI Security Baseline 1.0 for LLM Apps
AI that answers Security Questionnaires & RFP like you would
Simplifying OPSEC in cyber security investigations
Latest jobs in Cyber Security and Information Security.
Mindgard – Cyber Security Platform for AI