Skip to main content
Glama
tanevanwifferen

UseScraper MCP Server

UseScraper MCP 服务器

铁匠徽章这是一个基于 TypeScript 的 MCP 服务器,使用 UseScraper API 提供网页抓取功能。它提供了一个名为“scrape”的工具,可以从各种格式的网页中提取内容。

特征

工具

  • scrape从网页中提取内容

    • 参数

      • url (必填):要抓取的网页的 URL

      • format (可选):保存内容的格式(text、html、markdown)。默认值:markdown

      • advanced_proxy (可选):使用高级代理来规避机器人检测。默认值:false

      • extract_object (可选):指定要提取的数据的对象

Related MCP server: MCP-Typescribe

安装

通过 Smithery 安装

要通过Smithery自动为 Claude Desktop 安装 UseScraper:

npx -y @smithery/cli install usescraper-server --client claude

手动安装

  1. 克隆存储库:

    git clone https://github.com/your-repo/usescraper-server.git
    cd usescraper-server
  2. 安装依赖项:

    npm install
  3. 构建服务器:

    npm run build

配置

要与 Claude Desktop 一起使用,请添加服务器配置:

在 MacOS 上: ~/Library/Application Support/Claude/claude_desktop_config.json在 Windows 上: %APPDATA%/Claude/claude_desktop_config.json

{
  "mcpServers": {
    "usescraper-server": {
      "command": "node",
      "args": ["/path/to/usescraper-server/build/index.js"],
      "env": {
        "USESCRAPER_API_KEY": "your-api-key-here"
      }
    }
  }
}

/path/to/usescraper-server替换为服务器的实际路径,并将your-api-key-here替换为您的 UseScraper API 密钥。

用法

配置完成后,您可以通过 MCP 界面使用“scrape”工具。示例用法:

{
  "name": "scrape",
  "arguments": {
    "url": "https://example.com",
    "format": "markdown"
  }
}

发展

对于使用自动重建的开发:

npm run watch

调试

由于 MCP 服务器通过 stdio 进行通信,调试起来可能比较困难。我们推荐使用MCP Inspector ,它以包脚本的形式提供:

npm run inspector

检查器将提供一个 URL 来访问浏览器中的调试工具。

Available Tools

1 tool
scrapeC

Scrape content from a webpage using UseScraper API

ParametersJSON Schema
NameRequiredDescriptionDefault
urlYesURL to scrape
formatNoFormat to save crawled page content. Strongly recommended to keep as markdown for optimal AI processing (default: markdown)
advanced_proxyNoUse advanced proxy to circumvent bot detection (default: false)
extract_objectNoOptional object specifying data to extract

TDQS

C2.9/5.0
Behavior2/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

With no annotations, the description fails to disclose behavioral traits beyond the API name. It does not mention rate limits, authentication, potential failures, or what happens with the proxy option. The schema provides some param hints, but the description adds no behavioral context.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is a single sentence, concise and front-loaded. It is appropriately sized for a simple tool, though it could include a bit more detail without becoming verbose.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness2/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

Given the tool has 4 parameters, nested objects, and no output schema, the description is too minimal. It does not explain what is returned or how the tool behaves in edge cases, leaving significant gaps for an agent.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema coverage is 100% with each parameter described, so the baseline is 3. The description adds minimal extra meaning beyond stating the API, so it meets the baseline without enhancing understanding.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description identifies the action (scrape) and resource (webpage), naming the API used (UseScraper). However, 'scrape' is broad and could be more specific about what types of content or how it extracts, but without siblings it's sufficiently clear.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines2/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

No guidance is provided on when to use this tool versus alternatives or any exclusions. The description only states what it does, leaving the agent to infer usage context.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Tool Schema Changelog

Recent tool additions, removals, and schema changes observed during successful MCP inspections.

  1. 1 tool updatev0.1.0
    • First observedscrape

TDQS

B3.2/5.0

Scored across 1 tool

Disambiguation5/5

Only one tool exists, so there is no possibility of confusion with other tools.

Naming Consistency5/5

With a single tool, naming is trivially consistent.

Tool Count3/5

A single tool for web scraping is borderline; it might suffice for basic use but feels thin for a full-featured scraper server.

Completeness3/5

The single scrape tool covers the core functionality, but lacks additional features like selecting specific elements or handling different output formats, which are common in web scraping APIs.

Maintenance

ActivityInactive
ResponsivenessNo issues

Related MCP Connectors

Related MCP Servers

  • A
    license
    C
    quality
    C
    maintenance
    MCP Server enabling integration with Scrapezy to retrieve structured data from websites.
    1
    12
    13
    MIT
  • A
    license
    Not graded
    quality
    D
    maintenance
    An MCP server that enables LLMs to understand and work with TypeScript APIs they haven't been trained on by providing structured access to TypeScript type definitions and documentation.
    11
    46
    MIT
  • A
    license
    A
    quality
    D
    maintenance
    A TypeScript-based web scraping server built on the Model Context Protocol that offers multiple export formats, content extraction rules, and support for both static and dynamic (SPA) websites.
    7
    12
    1
    MIT
  • A
    license
    Not graded
    quality
    D
    maintenance
    A TypeScript-based MCP server that provides backend API handling and facilitates communication between microservices. Features an organized structure with controllers, routes, and models for easy extensibility and maintenance.
    208
    1
    MIT