Skip to main content
Glama
everford

Fetcher MCP

🚀 Fetcher MCP - сервер Playwright Headless Browser

Добро пожаловать в репозиторий Fetcher MCP GitHub! В этом репозитории размещен сервер MCP для загрузки содержимого веб-страниц с помощью браузера Playwright headless.

🧠 О нас

Fetcher MCP разработан для использования возможностей искусственного интеллекта для эффективного извлечения содержимого веб-страниц. Используя браузер Playwright headless, этот сервер может перемещаться по веб-страницам и с легкостью извлекать нужную информацию.

Related MCP server: Fetch MCP

🎯 Основные характеристики

🤖 Извлечение контента с помощью искусственного интеллекта
🔗 Интеграция драматурга
🚀 Быстро и эффективно
🌟 Простая установка и настройка

📚 Подробности репозитория

  • Имя : fetcher-mcp

  • Описание : MCP-сервер для загрузки содержимого веб-страниц с помощью браузера Playwright headless.

  • Темы : ИИ, MCP, Драматург

📦 Последний релиз

Последнюю версию сервера Fetcher MCP можно загрузить по следующей ссылке:
Скачать Fetcher MCP

:information_source: Примечание:

Предоставленная ссылка ведет прямо к файлу приложения. Пожалуйста, не забудьте запустить приложение после загрузки.

Если ссылка недоступна или не работает, вы можете проверить раздел «Релизы» этого репозитория на предмет альтернативных вариантов загрузки.

🚀 Начать

Чтобы начать использовать сервер Fetcher MCP для загрузки контента, выполните следующие простые шаги:

  1. Загрузите последнюю версию по ссылке выше.

  2. Разархивируйте загруженный файл в желаемое место.

  3. Запустите приложение.

  4. При необходимости настройте параметры сервера.

  5. Начните получать содержимое веб-страниц без усилий!

🌐 Дополнительные ресурсы

Для получения дополнительной информации, ресурсов или поддержки относительно сервера Fetcher MCP посетите официальный сайт по адресу https://github.com/everford/fetcher-mcp/releases .

📝 Правила внесения вклада

Мы приветствуем вклады, направленные на улучшение сервера Fetcher MCP и повышение его мощности и эффективности. Если у вас есть идеи, предложения или улучшения, отправьте запрос на извлечение, следуя нашим рекомендациям.

🙌 Присоединяйтесь к нашему сообществу

Общайтесь с другими разработчиками, делитесь идеями и будьте в курсе последних новостей, связанных с сервером Fetcher MCP, присоединившись к нашему сообществу:

👥 Канал Slack
🐦 Твиттер
📧 Рассылка новостей


🚀 Начните использовать сервер Fetcher MCP сегодня для бесперебойного извлечения содержимого веб-страниц с помощью возможностей на базе искусственного интеллекта. Легко извлекайте необходимую информацию с помощью интеграции браузера Playwright headless. Счастливого извлечения! 🌟


Помните, сервер Fetcher MCP упрощает процесс извлечения содержимого веб-страниц, делая его быстрее и эффективнее, чем когда-либо прежде. Загрузите последнюю версию сейчас и испытайте мощь ИИ и Playwright в действии. Удачной загрузки! 🚀

Available Tools

2 tools
fetch_urlC

Retrieve web page content from a specified URL

ParametersJSON Schema
NameRequiredDescriptionDefault
debugNoWhether to enable debug mode (showing browser window), overrides the --debug command line flag if specified
disableMediaNoWhether to disable media resources (images, stylesheets, fonts, media), default is true
extractContentNoWhether to intelligently extract the main content, default is true
maxLengthNoMaximum length of returned content (in characters), default is no limit
navigationTimeoutNoMaximum time to wait for additional navigation in milliseconds, default is 10000 (10 seconds)
returnHtmlNoWhether to return HTML content instead of Markdown, default is false
timeoutNoPage loading timeout in milliseconds, default is 30000 (30 seconds)
urlYesURL to fetch
waitForNavigationNoWhether to wait for additional navigation after initial page load (useful for sites with anti-bot verification), default is false
waitUntilNoSpecifies when navigation is considered complete, options: 'load', 'domcontentloaded', 'networkidle', 'commit', default is 'load'

TDQS

C2.9/5.0
Behavior2/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

With no annotations provided, the description carries the full burden of behavioral disclosure but only states the basic action. It fails to mention critical traits such as rate limits, authentication needs, potential for blocking or CAPTCHAs, error handling, or what 'retrieve' entails (e.g., using a headless browser, returning structured data). The description is too minimal for a tool with 10 parameters and complex web interactions.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is a single, efficient sentence that front-loads the core purpose without unnecessary words. It earns its place by clearly stating what the tool does, making it highly concise and well-structured for quick understanding.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness2/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

Given the tool's complexity (10 parameters, web scraping functionality) and lack of annotations and output schema, the description is incomplete. It doesn't address behavioral aspects, error cases, or return values, leaving significant gaps for an agent to understand how to use it effectively in real-world scenarios.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

The description adds no parameter-specific information beyond what the input schema provides. Since schema description coverage is 100%, with detailed descriptions for all 10 parameters, the baseline score of 3 is appropriate. The description doesn't compensate but doesn't need to, as the schema fully documents parameters like 'debug', 'extractContent', and 'waitUntil'.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description clearly states the tool's purpose as retrieving web page content from a URL, using specific verbs ('retrieve') and resources ('web page content', 'specified URL'). It distinguishes the core function but doesn't explicitly differentiate from the sibling tool 'fetch_urls', which appears to be a plural/multiple URL version.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines2/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description provides no guidance on when to use this tool versus alternatives like 'fetch_urls' or other web scraping methods. It lacks context about prerequisites, limitations, or typical use cases, leaving the agent with no usage direction beyond the basic purpose.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

fetch_urlsC

Retrieve web page content from multiple specified URLs

ParametersJSON Schema
NameRequiredDescriptionDefault
debugNoWhether to enable debug mode (showing browser window), overrides the --debug command line flag if specified
disableMediaNoWhether to disable media resources (images, stylesheets, fonts, media), default is true
extractContentNoWhether to intelligently extract the main content, default is true
maxLengthNoMaximum length of returned content (in characters), default is no limit
navigationTimeoutNoMaximum time to wait for additional navigation in milliseconds, default is 10000 (10 seconds)
returnHtmlNoWhether to return HTML content instead of Markdown, default is false
timeoutNoPage loading timeout in milliseconds, default is 30000 (30 seconds)
urlsYesArray of URLs to fetch
waitForNavigationNoWhether to wait for additional navigation after initial page load (useful for sites with anti-bot verification), default is false
waitUntilNoSpecifies when navigation is considered complete, options: 'load', 'domcontentloaded', 'networkidle', 'commit', default is 'load'

TDQS

C2.9/5.0
Behavior2/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

No annotations are provided, so the description carries the full burden of behavioral disclosure. It states the tool retrieves content but lacks details on critical behaviors: it doesn't mention authentication needs, rate limits, error handling, or what the output looks like (e.g., format, structure). For a tool with 10 parameters and no output schema, this is a significant gap in transparency.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is a single, efficient sentence: 'Retrieve web page content from multiple specified URLs.' It's front-loaded with the core purpose, has zero wasted words, and is appropriately sized for the tool's complexity. Every part of the sentence earns its place by clearly stating the action and scope.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness2/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

Given the tool's complexity (10 parameters, no annotations, no output schema), the description is incomplete. It doesn't address behavioral aspects like how content is returned, error cases, or performance constraints. While the schema covers parameters well, the description fails to provide necessary context for effective use, especially without annotations or output schema to fill gaps.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

The schema description coverage is 100%, meaning all parameters are well-documented in the input schema itself. The description doesn't add any semantic details beyond what's in the schema (e.g., it doesn't explain how 'urls' are processed or interactions between parameters). With high schema coverage, the baseline score of 3 is appropriate, as the description doesn't compensate but also doesn't detract.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description clearly states the tool's purpose: 'Retrieve web page content from multiple specified URLs.' It specifies the verb ('Retrieve'), resource ('web page content'), and scope ('multiple specified URLs'), which is specific and actionable. However, it doesn't explicitly distinguish this tool from its sibling 'fetch_url' (which presumably handles single URLs), missing full differentiation for a top score.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines2/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description provides no guidance on when to use this tool versus alternatives. It doesn't mention the sibling tool 'fetch_url' or explain scenarios where fetching multiple URLs is preferred over single ones. There's no context about prerequisites, limitations, or best practices, leaving the agent without usage direction.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Tool Schema Changelog

Recent tool additions, removals, and schema changes observed during successful MCP inspections.

  1. 2 tool updatesv1.0.0
    • First observedfetch_url
    • First observedfetch_urls

TDQS

C2.9/5.0

Scored across 2 tools

Disambiguation3/5

The two tools have overlapping purposes—both fetch web page content—but the descriptions clarify that one handles a single URL while the other handles multiple URLs. This distinction is clear enough to avoid misselection, but the core functionality is identical, leading to some ambiguity in why they are separate tools.

Naming Consistency5/5

The tool names follow a perfectly consistent verb_noun pattern with 'fetch_url' and 'fetch_urls', using snake_case throughout. The naming is predictable and clear, with no deviations in style or convention.

Tool Count2/5

With only two tools, the server feels under-scoped for a general-purpose 'Fetcher' domain. A single tool with parameters for single or multiple URLs could suffice, making the current count seem redundant and inefficient for typical agent workflows.

Completeness2/5

The tool surface is severely incomplete for web fetching; it lacks essential operations like handling HTTP methods (e.g., POST), managing headers, parsing content, or error handling. Agents will face dead ends when needing more than basic retrieval, causing frequent failures.

Related MCP Connectors

Related MCP Servers

  • A
    license
    A
    quality
    D
    maintenance
    An advanced web browsing server enabling headless browser interactions via a secure API, providing features like navigation, content extraction, element interaction, and screenshot capture.
    6
    27
    MIT
  • A
    license
    B
    quality
    F
    maintenance
    An MCP server that retrieves web page content using Playwright headless browser, capable of extracting main content and converting to Markdown format.
    3
    3,627 npm
    1,087
    MIT
  • A
    license
    B
    quality
    D
    maintenance
    A browser automation server providing Playwright capabilities for controlling web browsers, capturing screenshots, extracting content, and performing complex interactions through an MCP interface.
    6
    Apache 2.0