Home › Search: ai-safety

Search: ai-safety

Search MCP servers and agent skills by name, description, category or topic — 25 results.

MCP Server Maintained

jnMetaCode/shellward

AI Agent Security Middleware & MCP Server with 8-layer defense including prompt injection detection, DLP data flow tracking, command blocking, and PII detection. 7 MCP tools, zero dependencies.

128 TypeScript Updated 1mo ago Score 64
MCP Server Active

IgorGanapolsky/ThumbGate

MCP server that blocks AI coding agents from repeating mistakes — turns thumbs-up/down feedback into enforced pre-action gates via PreToolUse hooks. Install: `npx thumbgate`.

24 JavaScript Updated today Score 64
MCP Server Active

IgorGanapolsky/mcp-memory-gateway

Pre-action gates that prevent AI coding agents from repeating known mistakes. Captures explicit feedback, auto-promotes failures into prevention rules, and enforces them via hooks.

24 JavaScript Updated today Score 64
MCP Server Official Active

sekera-radim/impri

Human-in-the-loop approval inbox for AI agents. An agent submits a proposed action (send email, post comment, run a command) via `impri_push_action`, a human approves, rejects, or edits it from a web, mobile, or Slack/Discord/Telegram inbox, and the agent only proceeds on an approved decision. The gate is a data dependency, not a prompt. Full audit trail, self-hostable (MIT, Docker Compose). `npx

1 HTML Updated today Score 63
MCP Server Active

agentkitai/agentlens

Tamper-evident observability for AI agents: a SHA-256 hash-chained audit log with chain verification and signed export (EU AI Act Art. 12). Instrument any agent with zero code via `npx -y @agentlensai/mcp`; also ingests OpenTelemetry GenAI traces.

15 TypeScript Updated today Score 62
MCP Server Active

sint-ai/sint-protocol

Security-first MCP governance proxy (`sint-mcp`) with capability tokens, T0-T3 approval tiers, fail-closed execution, and tamper-evident audit receipts. Includes a separate `sint-scan` CLI for preflight MCP tool-risk audits.

11 TypeScript Updated today Score 61
MCP Server Stale

Govcraft/rust-docs-mcp-server

Provides up-to-date documentation context for a specific Rust crate to LLMs via an MCP tool, using semantic search (embeddings) and LLM summarization.

289 Rust Updated 8mo ago Score 58
MCP Server Active

AperionAI/shield

Local guardrail proxy for AI coding agents. Wraps any MCP server (stdio or Streamable HTTP) and blocks destructive tool calls — DROP TABLE, rm -rf, force-push — before they execute. MCP supply-chain protection: TOFU tool-catalog pinning against rug pulls, plus tool-description and tool-result scanning for tool poisoning and prompt injection. 51 starter rules, approval gates, audit logging. Single

6 Rust Updated 4d ago Score 58
MCP Server Active

kiro0x/five-mcp

LLM character consistency engine — generates structured JSON constraints from 4 multiple-choice questions about an AI's psychology. Drop the JSON into any LLM's system prompt to prevent persona drift; reduces inference cost from retries. 160,000 personality patterns; works with any LLM.

5 Python Updated 20d ago Score 58
MCP Server Active

lacs-project/sysknife

Security-hardened MCP server for Linux system administration via 189 typed actions instead of shell strings, with an Ed25519-signed hash-chain audit log, one-time TTL approval receipts, and automatic rollback. Works with Claude Code, Cursor, and Codex CLI.

4 Rust Updated yesterday Score 57
MCP Server Active

babyblueviper1/invinoveritas

Lightning-native AI reasoning, decisions, persistent memory, and agent marketplace for autonomous agents. Pay-per-use via Bitcoin Lightning. Register free — 250 starter sats. Agents earn sats selling services (seller keeps 95%), DM each other, and run autonomously. `npm install invinoveritas-mcp`

4 Python Updated today Score 57
MCP Server Active

bmdhodl/agent47

Runtime guardrails and incident read access for coding agents. Query AgentGuard traces, alerts, usage, costs, and budget health.

4 Python Updated 3d ago Score 57
MCP Server Active

behrensd/mcp-firewall

Deterministic security proxy (iptables for MCP) that intercepts tool calls, enforces YAML policies, scans for secret leakage, and logs everything. No AI, no cloud.

4 TypeScript Updated 19d ago Score 57
MCP Server Active

knowledgepa3/gia-mcp-server

Enterprise AI governance layer with 29 tools: MAI decision classification (Mandatory/Advisory/Informational), hash-chained forensic audit trails, human-in-the-loop gates, compliance mapping (NIST AI RMF, EU AI Act, ISO 42001), governed memory packs, and site reliability tools.

3 TypeScript Updated 2d ago Score 56
MCP Server Maintained

Acacian/aegis

Policy-based governance for AI agent tool calls. YAML policies, approval gates, risk assessment, and audit logging. Cross-platform: LangChain, OpenAI, Anthropic, MCP.

15 Python Updated 2mo ago Score 55
MCP Server Maintained

Chimera-Protocol/csl-core

Deterministic AI safety policy engine with Z3 formal verification. Write, verify, and enforce machine-verifiable constraints for AI agents via MCP.

15 Python Updated 2mo ago Score 55
MCP Server Active

loicfontaine-max/qorami-sdk

Check an email before an AI agent sends it: returns send / ask-a-human / block, with machine reason codes and prompt-injection detection.

2 Python Updated 29d ago Score 55
MCP Server Active

airblackbox/air-blackbox-mcp

EU AI Act compliance scanner for Python AI agents. Scans, analyzes, and remediates LangChain/CrewAI/AutoGen/OpenAI code across 6 articles with 10 tools including prompt injection detection, risk classification, and trust layer integration. The only MCP compliance server that generates fix code, not just findings.

2 Python Updated 18d ago Score 55
MCP Server Maintained

ark-forge/mcp-eu-ai-act

EU AI Act compliance scanner that detects regulatory violations in AI codebases with risk classification and remediation guidance.

11 Python Updated 1mo ago Score 54
Agent Skill Stale

frmoretto/clarity-gate

Epistemic quality verification for RAG systems

32 Python Updated 5mo ago Score 48
MCP Server Maintained

wei9072/aegis

AI-agent admission-control MCP server: validates file edits against Ring 0 syntax + Ring 0.5 structural-cost regression + workspace boundary (path / glob / shell-redirect / symlink). Negative-space framing — emits BLOCK / WARN / PASS verdicts, never coaches the agent.

1 Python Updated 2mo ago Score 46
MCP Server Maintained

cuttalo/depscope

Package Intelligence for AI agents. 22 tools across 17 ecosystems (npm/pypi/cargo/go/maven/nuget/rubygems/composer/pub/hex/swift/cocoapods/cpan/hackage/cran/conda/homebrew) — check health, vulnerabilities (OSV + CISA KEV + EPSS), typosquats, malicious flags, alternatives, known bugs, breaking changes, stack compatibility and error-to-fix. 31k+ packages, 2.2k+ CVEs enriched. Zero auth, MIT. Remote

1 TypeScript Updated 2mo ago Score 46
MCP Server Stale

sim-xia/blind-auditor

A zero-cost MCP server that forces AI to self-correct generation messages using prompt injection, independent self-audition and context isolation.

11 Python Updated 7mo ago Score 44
MCP Server Active

bluetieroperations-create/blackwall-mcp

Pre-action risk gate for AI agents. One `forecast` tool the agent calls before any irreversible action (send money, run SQL, delete data); returns a risk score (0–100), reversibility class, named red flags from 28 failure modes, and a gate: proceed / confirm / human-required.

0 JavaScript Updated 19d ago Score 40
MCP Server Archived

imran-siddique/agentos-mcp-server

Agent OS MCP server for AI agent governance with policy enforcement, code safety verification, multi-model hallucination detection, and immutable audit trails.

72 Python Updated 5mo ago Score 39