I spent 8 months trying to make LLMs Hack
I spent 8 months trying to make LLMs Hack
Hey HN! For that last 8 months I've been trying to make agents that can hack web applications to find vulnerabilities in them - An AI Security Tester. The system has 29 agents in total, a custom LLM Orchestration framework which works on the task-subtask architecture (old-school but works amazingly for my use case, and is pretty reliable) with custom agent calling mechanism. No Auo-Gen, Langchain and Crew AI - Everything custom built for pentesting. Each test runs in an isolated Kali linux environment (on AWS Fargate), where the agents have full access to the environment to undertake any step to hack the web application and find vulnerabilities. The agents have full access to the internet (through tavily) to search up and research content while conducting the test. After the test has been completed, which can take anywhere from 2-12 hours depending on the target, Peneterrer gives a full Vulnerability Management portal + A Pentest report completely generated by AI (sometimes 30+ pages long) You can test it out here - https://peneterrer.com/ Feedback appreciated!
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Incorrect prediction on native model
Similar products
November 2020: Four Months of Notado
Millionshort, 3 months later
After 9 months of work, we launched on Kickstarter
GPTCache – Redis for LLMs
prompttest – pytest for LLMs
Make your own end2end platform for LLMs in under 4 minutes
Your life visualized as 1080 months
My Isometric Voxel Engine 6 Months Later
My passion project for the last 6 months
My passion project for the last 6 months