Kilo-Kit
[](https://mseep.ai/app/vodailocz-kilo-kit-mcp)
<p align="center">
<img src="assets/social-preview.png" alt="Kilo-Kit social preview" width="100%">
</p>
<p align="center">
<a href="https://github.com/VoDaiLocz/kilo-kit-mcp/stargazers"><img src="https://img.shields.io/github/stars/VoDaiLocz/kilo-kit-mcp?style=for-the-badge&logo=github&label=Stars&color=18181b&labelColor=0f172a" alt="GitHub stars"></a>
<a href="https://github.com/VoDaiLocz/kilo-kit-mcp/commits/main"><img src="https://img.shields.io/github/last-commit/VoDaiLocz/kilo-kit-mcp?style=for-the-badge&logo=git&label=Last%20commit&color=22c55e&labelColor=0f172a" alt="Last commit"></a>
<a href="https://github.com/VoDaiLocz/kilo-kit-mcp/graphs/contributors"><img src="https://img.shields.io/github/contributors/VoDaiLocz/kilo-kit-mcp?style=for-the-badge&logo=github&label=Contributors&color=f97316&labelColor=0f172a" alt="Contributors"></a>
<a href="https://github.com/VoDaiLocz/kilo-kit-mcp/actions/workflows/publish.yml"><img src="https://img.shields.io/github/actions/workflow/status/VoDaiLocz/kilo-kit-mcp/publish.yml?style=for-the-badge&logo=githubactions&label=Publish&color=22c55e&labelColor=0f172a" alt="Publish workflow"></a>
<a href="https://www.npmjs.com/package/@vodailocz/kilo-kit-mcp"><img src="https://img.shields.io/npm/v/@vodailocz/kilo-kit-mcp?style=for-the-badge&logo=npm&label=npm&color=ef4444&labelColor=0f172a" alt="npm version"></a>
<a href="https://www.npmjs.com/package/@vodailocz/kilo-kit-mcp"><img src="https://img.shields.io/npm/dm/@vodailocz/kilo-kit-mcp?style=for-the-badge&logo=npm&label=downloads&color=0284c7&labelColor=0f172a" alt="npm downloads"></a>
<a href="https://github.com/VoDaiLocz/kilo-kit-mcp/blob/main/LICENSE"><img src="https://img.shields.io/github/license/VoDaiLocz/kilo-kit-mcp?style=for-the-badge&label=License&color=64748b&labelColor=0f172a" alt="License"></a>
</p>
<p align="center">
<img src="https://img.shields.io/badge/skills-180-06b6d4?style=for-the-badge&labelColor=0f172a" alt="180 skills">
<img src="https://img.shields.io/badge/tools-24%20MCP-10b981?style=for-the-badge&labelColor=0f172a" alt="24 MCP tools">
<img src="https://img.shields.io/badge/MCP-ready-14b8a6?style=for-the-badge&logo=modelcontextprotocol&labelColor=0f172a" alt="MCP ready">
<img src="https://img.shields.io/badge/Codex-ready-111827?style=for-the-badge&logo=openai&labelColor=0f172a" alt="Codex ready">
<img src="https://img.shields.io/badge/Node-%3E%3D20-339933?style=for-the-badge&logo=nodedotjs&labelColor=0f172a" alt="Node >=20">
<img src="https://img.shields.io/badge/TypeScript-5.9-3178c6?style=for-the-badge&logo=typescript&labelColor=0f172a" alt="TypeScript 5.9">
</p>
# Kilo-Kit: Autonomous Cognitive Flow & Quality Engine for AI Coding Agents
> **Version:** 1.9.1
> **Author:** Kilo-Kit Team
> **License:** Apache 2.0
Kilo-Kit is an agentic MCP runtime and curated 180-skill catalog designed to enforce grounded diagnosis, Tree-of-Thoughts architectural planning, adversarial red-teaming, 4D quality verification, and continuous SQLite self-improvement for AI coding assistants.
### ๐ง Core Architectural Pillars:
1. **Division of Labor (Cortex vs Limbs):** Kilo-Kit acts as the high-level cognitive brain (Tree of Thoughts, 5-Whys root cause analysis, adversarial stress-testing, context compaction) while host clients handle surgical I/O.
2. **Kilo-Sentinel Supervisor & Circuit Breaker:** Real-time middleware enforcing Pre-flight Grounding Locks (no editing unread files), loop tripwires (identical call and edit-thrashing detection), and SQLite trajectory logging (`katl_trajectories`).
3. **Triangulated Cognitive Synthesis & Low-Confidence Escalation:** Combines internal SQLite memory recall, GitHub 10k+ stars patterns, and ToT DAG benchmarking (`kilo_triangulate_research`), with automatic subagent delegation when confidence < 0.70.
4. **Fuzzy Skill & Alias Resolver:** Instant, resilient skill loading with support for aliases (`brainstorming`, `diagnose`, `playwright`, `clean-code`, `tdd`, `grounded-research-benchmark`).
5. **4D Quality Assurance & Playwright E2E Gate:** Validates Given-When-Then acceptance criteria, clean code interfaces, UI/UX aesthetics, and automated Playwright browser/DOM verification before work is marked complete.
---
## ๐๏ธ System Architecture & Division of Labor
Kilo-Kit enforces a strict architectural boundary between **High-Level Cognitive Reasoning (Cortex)** and **Surgical I/O Execution (Limbs)**:
```mermaid
flowchart TD
Clients["๐ฅ๏ธ Host AI Clients<br/>(Claude Code / Antigravity / Cursor / Gemini CLI)"]
Sentinel["๐ก๏ธ Kilo-Sentinel Supervisor & Circuit Breaker<br/>(Pre-flight Grounding Lock โข Loop Tripwire โข Step Budget)"]
subgraph Cortex["๐ง Kilo-Kit Cognitive Cortex (MCP Runtime)"]
direction TB
C4["๐๏ธ C4 5-Gate Lifecycle Controller"]
Engines["โ๏ธ 6 Cognitive Engines<br/>(ToT DAG โข Adversarial Grill โข 5-Whys โข Grounded Synthesis)"]
Skills["๐ 180 Curated Skills Catalog"]
Limbs["๐ ๏ธ Safe Execution Limbs<br/>(Atomic Write โข AST Edit โข Security Filtered Exec)"]
C4 --> Engines
C4 --> Skills
C4 --> Limbs
end
DB[("๐พ SQLite Atomic Memory<br/>(cognitive_triangulations โข katl_trajectories โข facts)")]
Clients <-->|MCP Protocol / stdio| Sentinel
Sentinel <--> Cortex
Cortex <--> DB
```
---
## ๐ C4 Cognitive Lifecycle & Low-Confidence Escalator
All agent tasks flow through a deterministic 5-Gate state machine. Unauthorized file mutations prior to Gate 3 approval are blocked at the server level:
```mermaid
flowchart LR
G1["<b>Gate 1: Grounded Probe</b><br/>โข Memory Recall<br/>โข Codebase Probe"]
G2["<b>Gate 2: Cognitive Reasoning</b><br/>โข 3-Option ToT DAG<br/>โข Adversarial Grill<br/>โข Low-Confidence Escalator"]
G3["<b>Gate 3: Approval</b><br/>โข Plan Locked<br/>โข Skills Injected"]
G4["<b>Gate 4: Execution</b><br/>โข Defense-in-Depth<br/>โข Sentinel Guarded"]
G5["<b>Gate 5: 4D QA</b><br/>โข Playwright E2E<br/>โข SQLite Reflection"]
G1 --> G2
G2 -->|Confidence >= 0.70| G3
G2 -.->|Confidence < 0.70<br/>Escalation| Subagent["๐ฌ Research Subagent<br/>(GitHub & Docs Sandbox)"]
Subagent -.-> G2
G3 --> G4
G4 --> G5
G4 -.->|3x Loop Detected| Breaker["๐ Circuit Breaker<br/>(Supervised Reset)"]
Breaker -.-> G1
```
---
## ๐ก๏ธ Protocol-Level Hard-Gate Enforcement
Traditional prompt rules (`.cursorrules`, `CLAUDE.md`) suffer from prompt drift. Kilo-Kit enforces safety via **JSON-RPC Interceptor Middleware**:
```mermaid
flowchart TD
Req["๐ค User Prompt"] --> Agent["๐ค Host AI Agent"]
Agent -->|1. Unauthorized Edit Attempt| GateCheck{"๐ก๏ธ Kilo-Sentinel<br/>Pre-Flight Lock"}
GateCheck -->|โ Uninitialized / Unread File| Blocked["๐ 403 Hard-Gate Blocked<br/>(Forces Planning First)"]
Blocked --> Plan["๐ง Gate 1 & 2: C4 Planning<br/>(kilo_orchestrate_task + kilo_triangulate_research)"]
Plan --> SaveDB[("๐พ Commit CoT to SQLite")]
SaveDB --> Approval["๐ค User Approves Plan"]
Approval -->|2. Authorized Execution| GateCheck
GateCheck -->|โ
State = READY| Exec["๐ Gate 4: Safe Execution<br/>(kilo_write_file / kilo_edit_file)"]
Exec --> Verify["โ
Gate 5: 4D Verification & SQLite Reflection"]
```
---
## ๐งฐ The 24 All-in-One MCP Tools Suite
Kilo-Kit provides a complete, self-contained execution and cognitive runtime:
| Category | Tool | Description |
| :--- | :--- | :--- |
| **Gating & Orchestration** | `kilo_orchestrate_task` | C4 closed-loop gate. Enforces brainstorming and cognitive steps before code mutation. |
| | `kilo_route_intent` | Routes intent to best workflow chains, task modes, and rules. |
| | `kilo_get_skill` | Loads curated `SKILL.md` workflows with token-safe truncation and session tracking. |
| | `kilo_search_skills` | High-precision semantic and keyword search across 180 skills. |
| | `kilo_memory_report` | Inspects persistent SQLite decisions, facts, and sessions. |
| | `kilo_remember_fact` | Pins immutable operational rules and architectural decisions into SQLite `memory_facts`. |
| | `kilo_record_reflection` | **Self-Improvement**: Persists reflections, correct/wrong paths, and lessons to SQLite. |
| | `kilo_route_report` | Reports route telemetry, top skills, workflows, scores, and conflict penalties. |
| | `kilo_validate_skills` | Validates entire skill catalog against the quality gate. |
| **Cognitive Reasoning** | `kilo_triangulate_research` | **Grounded Synthesis & Low-Confidence Escalator**: Combines SQLite memory + GitHub grounding + 3-option ToT DAG, atomically commits reasoning to SQLite, and triggers research escalation when confidence < 0.70. |
| | `kilo_think_step` | **Tree of Thoughts DAG**: Step-by-step reasoning, 3-option trade-off matrix & hypothesis branching. |
| | `kilo_grill_plan` | **Adversarial Red-Teaming**: Inversion, simplification, mobile touch & concurrency stress testing. |
| | `kilo_trace_root_cause` | **5-Whys Diagnostic Engine**: Recursive causal back-propagation with regression test scaffolding. |
| | `kilo_compact_context` | **Cognitive Compactor**: 40-70% token savings while locking invariants. |
| | `kilo_synthesize_skill` | **Self-Evolution**: Distills solved patterns into reusable skills. |
| **Sentinel & Supervision** | `kilo_sentinel_status` | **Supervisor Telemetry**: Inspects circuit breaker state, step budget, and grounded files list. |
| | `kilo_reset_circuit_breaker` | **Supervised Reset**: Resets tripped circuit breaker with root-cause justification. |
| | `kilo_benchmark_solution` | **Industry Benchmark**: Audits trajectory against GitHub standards and triggers re-planning. |
| **Safe Execution Suite** | `kilo_read_file` | Line slicing, size capping, and repository boundary enforcement. |
| | `kilo_search_files` | Glob pattern search across directory trees. |
| | `kilo_grep_code` | Line-by-line regex and substring search. |
| | `kilo_write_file` | Atomic write with **Protocol Hard-Gate**, clean-code smell audit, and secret detection. |
| | `kilo_edit_file` | Targeted search-and-replace with **JSON syntax & bracket balancing audit**. |
| | `kilo_run_command` | Defense-in-depth terminal execution with **security guardrails & command injection filtering**. |
---
## ๐ 180 Curated Skills Catalog Taxonomy
Skills are organized into 6 functional modules with instant alias mapping:
```
skills/
โโโ ๐๏ธ engineering/ (39 skills)
โ โโโ backend-development, codebase-design, api-patterns, database-design
โ โโโ nextjs-best-practices, react-patterns, tailwind-patterns, aspnet-core, better-auth
โโโ ๐งฉ problem-solving/ (24 skills)
โ โโโ sequential-thinking, root-cause-tracing, systematic-debugging
โ โโโ collision-zone-thinking, scale-game, simplification-cascades, inversion-exercise
โโโ ๐ productivity/ (33 skills)
โ โโโ brainstorming, spec-driven-development, tdd-workflow, code-review
โ โโโ grounded-research-benchmark, verification-before-completion, grill-me, subagent-driven-development
โโโ ๐ค agent-frameworks/ (26 skills)
โ โโโ workflow-state-machines, agent-memory, agentic-rag, multi-agent-orchestration
โ โโโ mcp-agent-patterns, code-agent-patterns, context-optimization
โโโ ๐ก๏ธ security/ (22 skills)
โ โโโ ai-guardrails, red-team-tactics, security-best-practices, vulnerability-scanner
โโโ โ๏ธ devops-cloud/ (36 skills)
โโโ devops, server-management, chrome-devtools, performance-profiling, render-deploy
```
**Fuzzy Alias Resolution:** Calling `kilo_get_skill("brainstorming")` automatically loads `productivity/brainstorming/SKILL.md`.
---
## โก Quick Start
### 1. Fast Setup (Recommended)
Install globally and automatically configure all detected AI clients (Cursor, Claude, Windsurf, Antigravity, Gemini):
```bash
npm install -g @vodailocz/kilo-kit-mcp
kilo-kit-init global
```
Verify system health:
```bash
kilo-kit-doctor
```
### 2. Zero-Install NPX (IDE Direct Integration)
Add Kilo-Kit directly to your client's MCP configuration without installing globally:
```json
{
"mcpServers": {
"kilo-kit": {
"command": "npx",
"args": ["-y", "@vodailocz/kilo-kit-mcp"]
}
}
}
```
*Supported in Cursor (`.cursor/mcp.json`), Claude Desktop (`claude_desktop_config.json`), Windsurf, and Antigravity / Gemini CLI.*
### 3. Team Repository Rollout
Bootstrap the Kilo-Kit C4 Cognitive Protocol into your project repository (`CLAUDE.md`, `AGENTS.md`, `GEMINI.md`):
```bash
kilo-kit-init init --client all
git add CLAUDE.md AGENTS.md GEMINI.md
git commit -m "chore: configure Kilo-Kit C4 protocol"
```
*Or use the global git alias: `git kilo-init`*
---
## ๐ Empirical Verification & Quality Benchmarks
| Metric | Without Kilo-Kit (Vanilla Agent) | With Kilo-Kit v1.9.1 | Verification Mechanism |
| :--- | :--- | :--- | :--- |
| **Silent Chained Tool Calls** | 65% on fast models (empty text outputs) | **0% (100% Enforced)** | Triple-Lock Schema + Sentinel Interceptor |
| **Ungrounded Code Mutations** | 42% of sessions (modifying unread files) | **0% (100% Blocked)** | Server-side Pre-flight Grounding Lock |
| **Context Window Longevity** | Degrades at >30k tokens | **Sustained >150k tokens** | `kilo_compact_context` (40โ70% token pruning) |
| **Silent Regression Rate** | 28% of PRs | **< 2%** | 4D QA (Playwright + Given-When-Then criteria) |
| **Tool Thrashing / Infinite Loops**| Common on complex bugs | **Terminated โค 3 loops** | Kilo-Sentinel Loop & Thrashing Tripwire |
| **Self-Healing & Reasoning Recall**| Zero across sessions | **100% SQLite Persistence**| `cognitive_triangulations` & `katl_trajectories` |
---
## ๐ Documentation Directory Index
For detailed specifications, protocol definitions, and developer guides, explore the [`docs/`](docs/) directory:
* ๐๏ธ [**Architecture Overview**](docs/architecture/ARCHITECTURE_DESIGN.md) - System topology, Cortex vs Limbs, and Memory schema.
* ๐ [**C4 Protocol Specification**](docs/protocols/C4_SPECIFICATION.md) - 5-Gate lifecycle state transitions and invariant rules.
* ๐งฐ [**MCP Tooling Reference**](docs/tools/TOOL_REFERENCE.md) - JSON schemas, error codes, and tool calling examples.
* ๐ [**Skills Taxonomy & Authoring**](docs/skills/SKILL_TAXONOMY.md) - Skill catalog guidelines and authoring standard.
* ๐ [**Benchmarks & Methodology**](docs/benchmarks/BENCHMARKS.md) - Token reduction experiments and SWE-bench alignment.
* ๐ก๏ธ [**Security & Guardrails**](docs/security/SECURITY_GUARDRAILS.md) - Circuit breaker, input sanitization, and blast-radius control.
---
## ๐งช Development, Testing & Verification
```bash
# Clone repository
git clone https://github.com/VoDaiLocz/KILO-KIT.git
cd KILO-KIT
npm install
# Run unit test suites (15/15 test files, 68/68 tests passing)
npm test
# Run full health diagnostics
npm run doctor
# Validate 100% of 180 skills against structural quality gates
node src/tools/validate-skill.js --all skills
```
---
## ๐ License
Distributed under the Apache 2.0 License. See [LICENSE](LICENSE) for more details.
TDQS
Scored across 24 tools
Several tools have overlapping meta-level purposes, especially kilo_orchestrate_task, kilo_route_intent, kilo_search_skills, and kilo_get_skill, which could cause misselection. The descriptions provide useful distinctions, but boundaries between workflow routing, skill discovery, and skill loading remain blurry.
All tools follow a consistent snake_case pattern with the same kilo_ prefix. The naming convention is predictable and easy to scan.
At 24 tools, the set is heavy for the server's apparent scope, sitting in the borderline range where each tool does not clearly earn its place. The domain is broad, but consolidation could reduce redundancy.
The surface covers skill discovery, loading, validation, synthesis, memory, routing, file operations, command execution, and several reasoning workflows. Minor gaps exist, such as no explicit skill deletion/update or memory fact deletion, but core workflows are well represented.