1. Conduid
  2. Marketplace
  3. #ai-safety
Tag

MCP servers tagged ai-safety

56 live MCP servers tagged "ai-safety", ranked by trust score. Tags come from package metadata and repository topics, so this list covers servers across every category that work with ai-safety.

56
live servers
46
avg trust
192
most stars
1

Template Repo

AndrewAltimit

Agent orchestration & security template featuring MCP tool building, agent2agent workflows, mechanistic interpretability on sleeper agents, and agent integration via CLI wrappers

★ 110updated 6 months agoUnlicense
83
2

Hegelion

Hmbown

Dialectical reasoning architecture for LLMs (Thesis → Antithesis → Synthesis)

★ 117updated 6 months agoMIT
80
3

Deterministic Agent Control Protocol

elliot35

Governance gateway for AI agents — bounded, auditable, session-aware control with MCP proxy, shell proxy & HTTP API. Works with Cursor, Claude Code, Codex, and any MCP-compatible…

★ 81updated 6 months agoMIT
77
4

mcp-guardian

eqtylab

MCP security scanner - detect prompt injection in tool descriptions

★ 192updated a year agoApache-2.0
73
5

Agentidentityprotocol

openagentidentityprotocol

Agent Identity Protocol - Zero-trust security layer for AI agents. Policy enforcement proxy for MCP with Human-in-the-Loop approval, DLP scanning, and audit logging.

★ 16updated 6 months agoApache-2.0
70
6

agentos-server

git+imran-siddique

[DEPRECATED] Moved to microsoft/agent-governance-toolkit

★ 67updated 6 months ago
67
7

Chimera-Protocol/csl-core

Chimera-Protocol

Deterministic safety layer for AI agents. Z3-verified policy enforcement.

★ 5updated 6 months agoApache-2.0
67
8

knowledgepa3/gia-mcp-server

knowledgepa3

MCP proxy for GIA Governance — connects Claude Desktop and Claude Code to the hosted GIA governance engine

★ 1updated 6 months agoNOASSERTION
67
9

Agentgate

agentkitai

API gateway for AI agents to access your personal data with human-in-the-loop write approval.

★ 6updated 6 months agoMIT
65
10

Qwed MCP

QWED-AI

MCP Server for QWED Verification - Use QWED verification tools in Claude Desktop, VS Code, and any MCP client

updated 6 months agoApache-2.0
65
11

Vibe Check MCP

kesslerio

Stop AI coding disasters before they cost you weeks. Real-time anti-pattern detection for vibe coders who love AI tools but need a safety net to avoid expensive overengineering tr…

★ 20updated 9 months agoNOASSERTION
63
12

Agentmesh MCP Proxy

Public Preview — Security proxy for MCP servers: The Firewall for Model Context Protocol

62
13

Read No Evil MCP

thekie

A secure email gateway MCP server that protects AI agents from prompt injection attacks hidden in emails.

★ 1updated 6 months agoApache-2.0
61
14

sim-xia/blind-auditor

Sim-xia

MCP tool for improving model coding quality by mandatory self-audition

★ 7updated 8 months agoMIT
53
15

Safeclaw

rawalrahul

Safe-by-default AI agent. Gates every action through Authensor before it executes.

★ 2updated 6 months agoMIT
52
16

Cordum

Cordum (cordum.io) is a platform-only control plane for autonomous AI Agents and external workers. It uses NATS for the bus, Redis for state and payload pointers, and CAP v2 wire…

52
17

Agentic AI

Agentic AI research papers, benchmarks, frameworks, and tools curated across 24 domains.

52
18

Reins

Runtime security for OpenClaw agents. Scan, fix, monitor.

52
19

Www Project Agent Memory Guard

OWASP Foundation web repository

52
20

Agent Safe Pipeline

Reference architecture for AI agents that propose actions but cannot authorize them — immutable intent capture, an independent Decionis policy verdict (ALLOW/ESCALATE/BLOCK), veri…

52
21

Node9 Proxy

The Execution Security Layer for the Agentic Era. Providing deterministic "Sudo" governance and audit logs for autonomous AI agents.

44
22

Pattern8

🎱 AI Agent Governance Framework — Constrain how AI Agents behave in your project. pip install pattern8

44
23

AI Driven Development

Practices, protocols, and skills for AI-driven software development. 18 skills + 1 Bash safety hook for Claude Code, Codex CLI, OpenCode, Cursor, Gemini CLI, Antigravity, and any…

44
24

Twin Github

Deterministic GitHub-shaped twin for agent testing — REST + MCP, SQLite-backed state.

42
25

Intercept

Control every MCP tool call your agent makes. Set budgets, approvals, and hard limits across MCP servers. So your agent can do its job without breaking things. Open Source

39
26

webcite-server

git+strategyconnect

WebCite MCP Server - Fact-check AI claims against authoritative sources. Works with any MCP-compatible agent or tool.

39
27

Avakill

🔪 Open-source safety firewall for AI agents. Intercepts tool calls before they execute, enforces YAML policies, and kills dangerous operations in real-time. Works with OpenAI, Ant…

39
28

Hardened Skills

200 AI agent skills, hardened with targeted behavioral guardrails. Free drop-in replacements.

39
29

IgorGanapolsky/ThumbGate

action gates via PreToolUse hooks. Install: `npx thumbgate`.

39
30

MCTS

MCTS (Model Context Threat Scanner) is a local-first security scanner for MCP servers -- static and live tool discovery, multiple analyzers, auditable risk scores, and JSON, SARIF…

39
31

Trustabl

Static analyzer for agent reliability.

39
32

LLM Differential Privacy Gateway

Noisegate: a differential privacy gateway that lets an untrusted LLM agent query sensitive data over MCP (Model Context Protocol), with a formal guarantee no individual's record c…

39
33

mcpwall

Deterministic security proxy for MCP tool calls. Blocks dangerous requests, redacts secrets from responses, catches prompt injection. No AI, no cloud, pure rules.

37
34

Phantom Secrets

MCP server for Phantom Secrets — lets AI coding tools manage API keys safely without ever seeing real values. 10 tools for Claude Code, Cursor, Windsurf, and Codex.

37
35

Quality Gate

Quality gate for MCP servers — compliance, security, and efficiency testing

37
36

Reagent

Governance layer for Claude Code — policy enforcement, hook-based safety gates, and audit logging for AI-assisted projects

37
37

Rea

Agentic governance layer for Claude Code — policy enforcement, hook-based safety gates, audit logging, and Codex-integrated adversarial review for AI-assisted projects

37
38

Altais

Modular, open-source MCP server providing comprehensive security analysis for AI coding agents.

37
39

Free AI Ops

Free MCP server for developers adding human review gates, prompt risk checks, FAQ review, and CRM note safety to AI tools.

37
40

Prism Verify

prism-verify — runtime LLM verifier (family-different, reasoning-stripped, multi-lens, signed Ed25519 receipts). Zero-prerequisite npx install via a verified binary launcher.

37
41

Metago Lifeform

MetaGO Agent Harness(智能体运行时控制层套件 · 驭智层)— 智能体运行时控制层,让智能体从工具升级为守规矩、会进化、可追溯、能闭环的生命体。39 技能 · 53 MCP tools · 8 公理 · Engine V2(KMWI 四层记忆 + 元进化五阶段 + 技能智能生成)· 决策锁四道关卡 · 全链路溯源。支持 7 大 AI 编程…

37
42

Governance

MCP Governance Server — wraps any MCP server with MAREF governance guardrails. Intercepts file writes and command execution to enforce policies via MAREF sidecar.

37
43

Border Patrol

A policy-enforcing MCP gateway: one MCP server that proxies all your others and blocks dangerous tool calls

37
44

Forge Plugin

Forge Plugin: Zero-dependency governance for Claude Code. 21 commands, 22 agents, 29 skills.

34
45

Map1

Deterministic identity for structured data

34
46

Sint Protocol

Open protocol and reference stack for governing AI agent actions in physical and safety-critical systems

34
47

Agent Guardrails

Merge gates and safety checks for AI coding agents. Works with Claude Code, Cursor, Windsurf, Codex via MCP. Detect scope violations, missing tests, and risks before merge.

34
48

sint-ai/sint-protocol

risk audits.

34
49

sidclawhq/mcp-guard

no SDK needed.

34
50

cuttalo/depscope

mcp`.

34
← Previous Page 1 of 2 Next →