← All tags

#ai-safety

75 posts

MCP Server 6

Still trusting results that AI generated, tested, and declared successful by itself? This Claude Code methodology adds evidence markers, aud...

Python MIT Updated 2w ago
MCP Server 4

The verification layer for autonomous agents — an independent, capital-aware verdict before an irreversible action (/review), a cryptographi...

Python Apache-2.0 Updated 2w ago
MCP Server 2

Runtime approval gates for AI agent tool calls. Intercept payments and emails before execution

Python Apache-2.0 Updated 3w ago
Claude Skill 18

Security scanner for AI agent skills. Detects prompt injection, data exfiltration, and malicious payloads before you install.

Python NOASSERTION Updated 2w ago
MCP Server 29

Lightweight AI safety middleware that protects humans by intercepting self-harm and criminal intent in LLM prompts. Features a 3-stage safet...

Python Apache-2.0 Updated 3w ago
MCP Server 11

MCP middleware that blocks dangerous AI agent actions using a simple YAML config

TypeScript Updated 1mo ago
Claude Skill 5

🔍 Discover + safety-vet Claude Code extensions before you install them - a 0-100 trust score for discovery + a 1-5 static-scan risk verdict...

Python MIT Updated 1mo ago
MCP Server 25

Cryptographically signed policies that constrain AI agent behavior. Same attestation primitives as cilock, applied to agent execution: ident...

Go Apache-2.0 Updated 2w ago
MCP Server 11

Stop AI coding agents from leaking your API keys. Local proxy + MCP that swaps real secrets for phm_ tokens — works with Claude Code, Cursor...

Rust MIT Updated 4w ago
MCP Server 25

AI coding safety CLI for vibe coding workflows. Checkpoints, undo, anchors, MCP, and secret protection for Claude Code, Cursor, Codex, and O...

Python MIT Updated 1mo ago
MCP Server 12

Policy-gated, durable, audited execution for AI agent tool calls. Every action gets a five-verdict policy check, a crash-safe checkpoint, an...

Python NOASSERTION Updated 3w ago
MCP Server 9

Bilingual hands-on roadmap for production-aware AI agents: MCP, memory, RAG, workflows, evaluation, safety, and agent colonies.

Python MIT Updated 4w ago