values-firewall
Check any proposed action against a declared policy charter and get back an auditable, advisory receipt. Two layers: deterministic hard-rule and pattern matching (no LLM), plus an LLM values assessment that cites the exact principles it applied. Returns a recommendation of allow, flag, or gate-to-human, never a hard approved/denied. Advisory and read-only: it recommends, it never blocks, executes, or authorizes, and it writes nothing. Fail-safe by design (if the values layer is unavailable it degrades to at least "flag", never a silent "allow"). Public demo runs against an open Article-14 demo charter.
First seen 2 Oct 2026. Evidence as of 5 Oct 2026.
Tools
No tool list captured yet.
Change history
No changes since the first observation. The first snapshot is the baseline.
| Source | Listing | First seen | Last seen | Versions |
|---|---|---|---|---|
| Smithery | HUMANIFIED/values-firewall | 2 Oct 2026 | 5 Oct 2026 | 1 |