HTTrack MCP Server
Click on "Deploy Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@HTTrack MCP Servermirror example.com into project example-site"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
HTTrack MCP Server
A FastMCP (Model Context Protocol) server that wraps the classic HTTrack website copier, so AI agents can mirror websites into a shared archive and then browse, read, and search the downloaded content — all over MCP.
Copyright (C) 2026 Tech Ventures VCC. Licensed under the GNU Affero General Public License v3.0 (see LICENSE).
Why
HTTrack is battle-tested for offline website copying, but it has no scripting API — only a
GUI and a CLI. This server makes it agent-native: an LLM agent can call mirror_site to
start a download in the background, poll mirror_status, and later answer questions from
the archive via catalog, list_files, read_file, and search_files.
Related MCP server: webserver-mcp
Tools
Tool | Purpose |
| Start a background mirror into the archive ( |
| Job status ( |
| All projects: short name, short description (auto-derived from each site's |
| Directory listing inside a project (name, type, size, modified). |
| Read one text file (HTML/JS/CSS/txt/json/md). Refuses binaries; caps content at 256 KB to keep agent context sane. |
| Case-insensitive substring search across a project's text files, capped results. |
Quick start
Prerequisites: Docker.
docker build -t httrack-mcp-server .
docker run -d --name httrack-mcp --restart unless-stopped \
-p 9050:8000 \
-v /path/to/your/archive:/data:rw \
httrack-mcpThe container exposes the MCP server over streamable HTTP at http://<host>:9050/mcp.
Register in an MCP client
Example for MCPHub-style gateways (streamable-http):
{
"httrack": {
"type": "streamable-http",
"url": "http://192.168.1.35:9050/mcp",
"enabled": true,
"description": "HTTrack site mirroring + archive browsing",
"owner": "admin"
}
}Clients that speak streamable HTTP directly can connect to /mcp with no extra auth
(nobody is authenticated; run this on a trusted network, or put it behind your gateway's
auth). See HELP.md for full usage, examples, and troubleshooting.
Concurrency model
One
httrackprocess per project, isolated by its own output directory — no shared state between projects, so multiple agents can mirror different sites in parallel.mirror_siteguards against double-mirroring the same project (in-process job table + a/procscan for httrack processes writing to that project directory).
CI
GitHub Actions runs on every PR and push: Python syntax check of server.py,
Dockerfile sanity assertions, and a Docker image build.
License
This program is free software: you can redistribute it and/or modify it under the terms of the GNU Affero General Public License as published by the Free Software Foundation, either version 3 of the License, or (at your option) any later version.
If you run a modified version of this server as a network service, AGPL §13 requires you to offer your users the corresponding source of that modified version.
Copyright (C) 2026 Tech Ventures VCC. HTTrack itself is GPL-licensed by Xavier Roche and contributors; this project only shells out to it and is not affiliated with it.
This server cannot be deployed
Maintenance
Related MCP Connectors
Turn any public website into an MCP server for agents to search, read and navigate.
Scrape, crawl and search the web for AI agents via MCP.
- hiveWikiOAuthai.hivewiki
Shared project wiki for AI agents: read and write pages, next actions, and activity logs over MCP.
Shared copies of public web pages for AI agents. Search stored pages or fetch a URL.
Related MCP Servers
- AlicenseNot gradedqualityCmaintenanceEnables AI agents to clone entire websites, download files, manage authentication sessions, and analyze site information with support for JavaScript-heavy SPAs and dynamic content.9Apache 2.0
- FlicenseNot gradedqualityDmaintenanceEnables AI agents to edit and serve a static website via natural language, providing file management tools over MCP and HTTP hosting.-
- AlicenseCqualityCmaintenanceMCP server for safely reading public URLs for AI agents, providing tools to fetch, extract, cache, and inspect web content as evidence.15MIT
- AlicenseNot gradedqualityBmaintenanceEnables AI agents to perform web searches with full content retrieval and multi-engine provenance, including trust scoring and local corpus persistence, via MCP integration.1 npm2Apache 2.0