I

I built a runtime governance layer for LLMs. Can you break it?

Hacker News

I built a runtime governance layer for LLMs. Can you break it?

I’ve spent the last year building SAFi, an open-source cognitive architecture that wraps around AI models (GPT, Claude, etc.) to enforce alignment with human values. Safi is a "System 2" architecture inspired by classical philosophy. It separates the generation from the decision: The Intellect: proposes a draft. The Will: decides to block or approve the drafts. The Conscience: audits the drafts based on set core values The Spirit: An EMA (Exponential Moving Average) vector that tracks "Ethical Drift" over time and injects course-correction into the context window. The Challenge: I want to see if this architecture actually holds up. I’ve set up a demo with a few agents. I want you to try to jailbreak them. Repo: https://github.com/jnamaya/SAFi Demo: https://safi.selfalignmentframework.com/ Homepage: https://selfalignmentframework.com/ Safi is licensed under GPLv3.

Share card

Actual performance

1points
Did not reach leaderboard

Launch Intel predictions

Analyze your own launch →
Product HuntOn track for Day 1 leaderboard · Strong signals: agents, agent, claude · Missing: mac, macos, cursor
88%88% predicted probability of success on Product Hunt, based on ML models trained on real launch data.
best fitHighest predicted score across all platforms for this description.
AppSumoMay struggle as an AppSumo deal · Missing: plus, platform, intuitive
48%48% predicted probability of success on AppSumo, based on ML models trained on real launch data.
Indie HackersIH features products with proven revenue · Strong signals: para · Missing: supports, reddit linkedin, podcasting
41%41% predicted probability of success on Indie Hackers, based on ML models trained on real launch data.
TrustMRRLess likely to generate early MRR · Strong signals: para · Missing: mobile apps, ios, personal
32%32% predicted probability of success on TrustMRR, based on ML models trained on real launch data.
Hacker NewsMay not resonate with HN audience · Strong signals: ide, io · Missing: https docs, excited, just released
30%30% predicted probability of success on Hacker News, based on ML models trained on real launch data.
nativeThis product was originally launched on this platform.
Acquire.comPre-revenue stage for this audience · Missing: arr, mrr, revenue
17%17% predicted probability of success on Acquire.com, based on ML models trained on real launch data.
BetaListMay not resonate with beta-testers · Missing: web3, chat, crypto
0%0% predicted probability of success on BetaList, based on ML models trained on real launch data.

Correct prediction on native model

Similar products

Preloop
Preloop79%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

The MCP Governance Layer

Product Hunt+121Open Source
SA
SAFi, a Governance Engine for LLMs30%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

SAFi, a Governance Engine for LLMs

Hacker News2
Ge
GenOps AI – OSS (OpenTelemetry) runtime governance for AI workloads35%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

GenOps AI – OSS (OpenTelemetry) runtime governance for AI workloads

Hacker News3
Ed
Edictum – Runtime governance for LLM agent tool calls43%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Edictum – Runtime governance for LLM agent tool calls

Hacker News2
To
Towards AI Governance31%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Towards AI Governance

Hacker News1
A
A whitepaper on runtime governance for autonomous AI systems22%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

A whitepaper on runtime governance for autonomous AI systems

Hacker News1
SilkFlo
SilkFlo54%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

AI Governance platform

Indie Hackers1$3,000/moai
A
A container runtime implemented in x86_64 assembly55%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

A container runtime implemented in x86_64 assembly

Hacker News3
Su
Subverting Go's Runtime System47%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Subverting Go's Runtime System

Hacker News2
To
Toggle Methods and Endpoints at Runtime48%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Toggle Methods and Endpoints at Runtime

Hacker News2