← All tags

#ai-safety

81 posts

MCP Server 6

Proof-backed AI agent for checking suspicious job posts, recruiter messages, and apply links, now live with case study and pilot intake.

TypeScript Updated 1mo ago
MCP Server 8

MCP servers expose tools with no information about what they actually do at runtime. mcpsafetywarden sits between your agent and any MCP ser...

Python NOASSERTION Updated 2mos ago
MCP Server 4

MCP & Claude Code security scanner — threat-models plugins, MCP servers, hooks, skills & connectors with an LLM before you trust them. Catch...

Go Apache-2.0 Updated 2mos ago
MCP Server 11

Glass Box Framework — runtime constitutional verification for AI answers. Trust Cards with claim-level reasoning chains, formal ECS scoring,...

TypeScript Apache-2.0 Updated 3mos ago
MCP Server 4

Runtime artifact existence & freshness verification for AI agent completion claims — a lightweight, zero-LLM MCP gate that source-binds 'don...

TypeScript MIT Updated 2mos ago
MCP Server 6

A safer MySQL CLI for AI coding agents: connection profiles, SSH tunnels, and automatic sensitive-data masking before query output reaches C...

Go MIT Updated 1mo ago
Claude Skill 41

🛡️ A curated list of resources on agent skills security: attacks, defenses, frameworks, and benchmarks for securing AI agent tool use and sk...

Updated 2mos ago
MCP Server 9

Open-source prompt injection detector — 5 layers, 91.7% F1, ~27ms, offline, Apache 2.0

Python Apache-2.0 Updated 3mos ago
MCP Server 15

Deterministic policy language for AI agents. Z3 + TLA+ dual-engine formal verification. Runtime enforcement <1ms.

Python Apache-2.0 Updated 3mos ago
MCP Server 7

Pre-execution policy engine for AI agents. Every tool call checked before execution.

TypeScript NOASSERTION Updated 4mos ago