Fi

Firewall for LLMs–Guard Against Prompt Injection, PII Leakage, Toxicity

Hacker News

Firewall for LLMs–Guard Against Prompt Injection, PII Leakage, Toxicity

Hey HN, We're building Aegis, a firewall for LLMs: a guard against adversarial attacks, prompt injections, toxic language, PII leakage, etc. One of the primary concerns entwined with building LLM applications is the chance of attackers subverting the model’s original instructions via untrusted user input, which unlike in SQL injection attacks, can’t be easily sanitized. (See https://greshake.github.io/ for the mildest such instance.) Because the consequences are dire, we feel it’s better to err on the side of caution, with something mutli-pass like Aegis, which consists of a lexical similarity check, a semantic similarity check, and a final pass through an ML model. We'd love for you to check it out—see if you can prompt inject it!, and give any suggestions/thoughts on how we could improve it: https://github.com/automorphic-ai/aegis . If you want to play around with it without creating an account, try the playground: https://automorphic.ai/playground . If you're interested in or need help using Aegis, have ideas, or want to contribute, join our Discord ( https://discord.com/invite/E8y4NcNeBe ), or feel free to reach out at founders@automorphic.ai. Excited to hear your feedback! Repository: https://github.com/automorphic-ai/aegis Playground: https://automorphic.ai/playground

Share card

Actual performance

17points
7comments
Made the leaderboard

Launch Intel predictions

Analyze your own launch →
Indie HackersFits the IH revenue-focused audience · Missing: supports, reddit linkedin, podcasting
60%60% predicted probability of success on Indie Hackers, based on ML models trained on real launch data.
best fitHighest predicted score across all platforms for this description.
Product HuntOn track for Day 1 leaderboard · Strong signals: model, user, using · Missing: mac, agents, macos
59%59% predicted probability of success on Product Hunt, based on ML models trained on real launch data.
Hacker NewsMay not resonate with HN audience · Strong signals: excited, ide, io · Missing: https docs, just released, exist
47%47% predicted probability of success on Hacker News, based on ML models trained on real launch data.
nativeThis product was originally launched on this platform.
TrustMRRLess likely to generate early MRR · Missing: mobile apps, ios, personal
43%43% predicted probability of success on TrustMRR, based on ML models trained on real launch data.
AppSumoMay struggle as an AppSumo deal · Missing: plus, platform, intuitive
43%43% predicted probability of success on AppSumo, based on ML models trained on real launch data.
Acquire.comPre-revenue stage for this audience · Missing: arr, mrr, revenue
13%13% predicted probability of success on Acquire.com, based on ML models trained on real launch data.
BetaListMay not resonate with beta-testers · Missing: web3, chat, crypto
0%0% predicted probability of success on BetaList, based on ML models trained on real launch data.

Incorrect prediction on native model

Similar products

Op
Open-Source Firewall for LLMs40%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Open-Source Firewall for LLMs

Hacker News4
Deckrd
Deckrd35%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

The Firewall for Humanity.

Indie Hackerscommitment-full-time
Bu
Built firewall for LLMs after prompt injection bypass GPT-4s guardrails36%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Built firewall for LLMs after prompt injection bypass GPT-4s guardrails

Hacker News1
UA
UAC Guard53%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

UAC Guard

Hacker News2
sn
sn00p – poc app firewall31%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

sn00p – poc app firewall

Hacker News2
Wardstone
Wardstone27%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Guard your AI from prompt attacks, harmful content, and PII

Product Hunt+2
A
A Firewall for AI agents with auditing32%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

A Firewall for AI agents with auditing

Hacker News4
Tr
Trajeckt, a Firewall for AI Agents30%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Trajeckt, a Firewall for AI Agents

Hacker News2
HOL Guard
HOL Guard20%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Local-first runtime firewall for AI coding agents

Indie Hackersrevenue-model-free
HOL Guard
HOL Guard9%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Local-first runtime firewall for AI coding agents

Indie Hackersfunding-self