A safety layer for AI coding agents. CLAUDE.md/AGENTS.md generator, MCP runtime guardrail, pre-commit hook, GitHub Action.
#ai-safety
81 posts
Agent skill for deepfake detection & media safety — detect AI-generated audio, images, and video with Resemble AI
Merge gates and safety checks for AI coding agents. Works with Claude Code, Cursor, Windsurf, Codex via MCP. Detect scope violations, missin...
AI Constraint Engine — enforces CLAUDE.md, .cursorrules, AGENTS.md rules as laws. 51 MCP tools, 991 tests. Official MCP Registry. npx speclo...
Enforce zero-trust rules for AI agents to prevent hallucinations, unsafe actions, and policy bypasses
Hands-on study companion for the Claude Certified Architect - Foundations (CCA-F) exam.
Real-time trustworthiness evaluation and safety interception for AI agents. Semantic analysis, safe alternative suggestions, multi-step atta...
Agentic security control plane for MCP and AI agent tool calls. MCP-native policy gateway with topology discovery and audit.
Dialectical reasoning architecture for LLMs (Thesis → Antithesis → Synthesis)
Security enforcement plugin for Claude Code. Blocks dangerous commands, audits every tool call, detects prompt injection.
LikenessGuard is an open-source reference project for pre-generation consent enforcement in AI image generation. V2 demonstrates the concept...
Your AI agent just burned $200. AgentGuard stops it at $5. Runtime cost guardrails for AI agents — budget enforcement, loop detection, kill...