ai-guardrails
Click on "Deploy Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@ai-guardrailsvalidate this observation: revenue is marked VERIFIED"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
🛡️ AI Guardrails Protocol (v0.1.0)
A small, free, MIT-licensed Python module built around one fundamental rule for autonomous AI agents: A CLAIM IS NOT EVIDENCE.
Why this exists
Most autonomous agent frameworks allow an LLM to declare its own completion and success. This self-evaluative hallucination is unsafe. This prototype provides a deterministic gate: an agent cannot mark its own observation as VERIFIED without an independent external verifier.
Attempting to self-verify raises an immediate GuardrailViolation exception instead of passing silently.
Usage Example
from guardrails import TruthBoundaryEnforcer, GuardrailViolation
# 1. Unverified observation passes cleanly into the state graph:
valid_record = {
"entity": "github_stats",
"state": "OBSERVED",
"verification_status": "UNVERIFIED"
}
TruthBoundaryEnforcer.validate_observation(valid_record) # OK
# 2. Self-proclaimed VERIFIED observation is blocked blocked:
hallucinated_record = {
"entity": "revenue",
"state": "OBSERVED",
"verification_status": "VERIFIED" # Missing independent_verifier!
}
try:
TruthBoundaryEnforcer.validate_observation(hallucinated_record)
except GuardrailViolation as e:
print(f"Blocked: {e}")MCP Server Configuration
An entry point for Model Context Protocol is included in mcp_server.py. You can configure it in Claude Desktop by adding this to your claude_desktop_config.json:
{
"mcpServers": {
"ai-guardrails": {
"command": "python3",
"args": ["/absolute/path/to/ai-guardrails-protocol/mcp_server.py"]
}
}
}Replace /absolute/path/to/ai-guardrails-protocol with the path to your own clone.
Current Status & Limitations
Prototype: This is a minimalist reference implementation.
No Paid Version: The code is completely free and open source under MIT.
Warranty: Provided as-is without warranty.
License
MIT License. Copyright (c) 2026 Yusuf Sen / PHI-BRAIN Systems.
This server cannot be deployed
Maintenance
Related MCP Connectors
Evidence-backed x402 web verification for AI agents, with auditable decisions for every condition.
Tamper-evident proof creation and verification for AI agents via MCP, A2A, and REST.
Verifier-grounded AI promotion gates, disposable report cards, and signed PASS/HOLD/BLOCK receipts.
Verify a source or provider before an AI agent trusts it. Evidence only; unknown stays unknown.
Related MCP Servers
- AlicenseNot gradedqualityCmaintenanceA deterministic verification gate for MCP clients that independently checks model outputs against evidence, contradictions, calibration, and provenance without relying on LLM self-assessment.1MIT
- FlicenseAqualityCmaintenanceEnables agents to verify claims against a real falsification ledger, querying pre-registered refutation thresholds, recorded verdicts, and verifier drift metrics instead of guessing.519 PyPI-
- FlicenseNot gradedqualityBmaintenanceEnables AI agents to verify trust credentials, check revocation status, and fingerprint MCP tool surfaces to detect poisoning or drift.1-
- AlicenseNot gradedqualityAmaintenanceEnables AI agents to verify claims against evidence through MCP tools, returning approved, rejected, or needs_review verdicts with replayable receipts.MIT