mcp-zyte
Provides tools for fetching web page content, AI-powered structured data extraction (e.g. product and article data), and capturing screenshots via the Zyte API.
Click on "Deploy Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@mcp-zyteExtract product info from https://example.com/item/123"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
mcp-zyte
Zyte API MCP — unified web fetch + AI extraction (zyte.com)
Part of Pipeworx — an MCP gateway connecting AI agents to 1576+ live data sources.
Tools
Tool | Description |
| Fetch a web page through Zyte API and return its content. By default returns the raw HTTP response body (decoded to text). For JS-heavy sites that need a real browser, pass render:true to get browser-rendered HTML instead. Example: zyte_fetch({ url: "https://example.com", _apiKey: "your-key" }) |
| AI-powered automatic extraction of structured data from a web page. Set type to 'product' for e-commerce product pages (name, price, currency, images, SKU, availability...) or 'article' for news/blog pages (headline, author, date, body text...). Example: zyte_extract({ url: "https://shop.example.com/item/123", type: "product", _apiKey: "your-key" }) |
| Capture a screenshot of a web page via Zyte API (rendered in a headless browser). Returns metadata about the base64-encoded PNG (its length) rather than inlining the full blob. Example: zyte_screenshot({ url: "https://example.com", _apiKey: "your-key" }) |
Related MCP server: ScrapeUnblocker MCP Server
Quick Start
Add to your MCP client (Claude Desktop, Cursor, Windsurf, etc.):
{
"mcpServers": {
"zyte": {
"url": "https://gateway.pipeworx.io/zyte/mcp"
}
}
}What this endpoint actually serves
tools/list at https://gateway.pipeworx.io/zyte/mcp returns the tools in the table
above plus the shared Pipeworx meta-tools — ask_pipeworx,
discover_tools, search_within, remember/recall and the rest of the
gateway-wide set. So the tool count you see is larger than this table: a
single-pack endpoint currently lists roughly 30 shared tools alongside the
pack's own. The connection's initialize response states its exact scope, and
is the authoritative answer for a given day.
This is deliberate, not multiplexing by accident. The meta-tools are what let a
scoped connection answer a question this pack does not cover — via
ask_pipeworx, which routes across the whole catalog — without you adding a
second MCP server. There is currently no way to mount a pack endpoint without
them; if the extra schemas cost you more context than the routing is worth,
connect to the full gateway once rather than to several pack endpoints.
Or connect to the full Pipeworx gateway to get every pack's tools listed directly, instead of just this one's:
{
"mcpServers": {
"pipeworx": {
"url": "https://gateway.pipeworx.io/mcp"
}
}
}Both URLs reach the same gateway and the same 1576+ data sources. The
only difference is which pack's tools are listed directly; ask_pipeworx
reaches all of them from either one.
No MCP client? Call it over HTTP
This pack takes your own API key (_apiKey) — we don't front one for it, so there's no curl here that would run without it. Inspect any tool: GET https://gateway.pipeworx.io/v1/tools/zyte_fetch. Find one: POST https://gateway.pipeworx.io/v1/tools/search_packs with {"query":"..."}.
Standalone (no gateway account)
This package also runs as a local stdio MCP server — no Pipeworx account, no gateway round-trip:
{
"mcpServers": {
"zyte": {
"command": "npx",
"args": ["-y", "@pipeworx/mcp-zyte"]
}
}
}Or run it directly to confirm it starts:
npx -y @pipeworx/mcp-zyteIt speaks MCP over stdin/stdout and answers initialize/tools/list/tools/call
for only this pack's tools — none of the shared meta-tools the gateway
connection above adds. Same source, same tools, no ask_pipeworx routing.
Using with ask_pipeworx
Instead of calling tools directly, you can ask questions in plain English — this works on the pack endpoint above as well as on the full gateway:
ask_pipeworx({ question: "your question about Zyte data" })The gateway picks the right tool and fills the arguments automatically.
More
License
MIT
This server cannot be deployed
Maintenance
Related MCP Connectors
Enable language models to perform advanced AI-powered web scraping with enterprise-grade reliabili…
Fetch and process content from specified URLs & sources using the Oxylabs Web Scraper API.
Crawl, scrape, search the web, and automate browsers at scale with anti-bot bypass.
Fetch pages as markdown, search web and news, extract structured data. For AI agents.
Related MCP Servers
- AlicenseDqualityAmaintenanceInteract with WebScraping.AI API for web data extraction and scraping731 npm43MIT
- AlicenseAqualityAmaintenanceEnables fetching any web page's HTML by bypassing anti-bot protection, and also provides AI-parsed structured data and Google search results.4104 npm1MIT
- AlicenseAqualityDmaintenanceEnables AI agents to browse the web, bypass anti-bots, render JavaScript, take screenshots, and perform structured data extraction using the ScrapeOps Proxy API.35 npmMIT
- FlicenseNot gradedqualityBmaintenanceEnables AI agents to operate a real, anti-bot-aware browser through Zyte API for navigating pages, clicking, typing, scrolling, taking screenshots, searching, and extracting structured data, with automatic proxy rotation and ban avoidance.9 npm-