Beating GPT5.5-xhigh for Coding agent security with SLMs and IRM
Beating GPT5.5-xhigh for Coding agent security with SLMs and IRM
Coding agents craft arbitrary code so securing them is more complicated than red-teaming. We post trained a cyber-security small llm, changed how it reasons and supplemented our controls using program analysis techniques such as inline reference monitoring to outperform GPT5.5-xhigh on hard benchmarks like LinuxArena and SleightBench. Free product available at harden.run and full benchmarks in the blog post.
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Incorrect prediction on native model
Similar products
Bestie, a coding agent that respects you
Bailout – The coding agent meant to be deleted
Premortem, a coding-agent-powered airplane blackbox
Website Coding Agent
Hazzel – terminal coding agent, BYOK, undo-everything
Neurogrid terminal coding agent
The coding agent that asks before it builds
The coding-agent harness you can make your own
zot – Yet another coding agent harness
Zot – Yet another coding agent harness