UseScraper MCP Server
UseScraper MCP 服务器
这是一个基于 TypeScript 的 MCP 服务器,使用 UseScraper API 提供网页抓取功能。它提供了一个名为“scrape”的工具,可以从各种格式的网页中提取内容。
特征
工具
scrape从网页中提取内容参数:
url(必填):要抓取的网页的 URLformat(可选):保存内容的格式(text、html、markdown)。默认值:markdownadvanced_proxy(可选):使用高级代理来规避机器人检测。默认值:falseextract_object(可选):指定要提取的数据的对象
Related MCP server: MCP-Typescribe
安装
通过 Smithery 安装
要通过Smithery自动为 Claude Desktop 安装 UseScraper:
npx -y @smithery/cli install usescraper-server --client claude手动安装
克隆存储库:
git clone https://github.com/your-repo/usescraper-server.git cd usescraper-server安装依赖项:
npm install构建服务器:
npm run build
配置
要与 Claude Desktop 一起使用,请添加服务器配置:
在 MacOS 上: ~/Library/Application Support/Claude/claude_desktop_config.json在 Windows 上: %APPDATA%/Claude/claude_desktop_config.json
{
"mcpServers": {
"usescraper-server": {
"command": "node",
"args": ["/path/to/usescraper-server/build/index.js"],
"env": {
"USESCRAPER_API_KEY": "your-api-key-here"
}
}
}
}将/path/to/usescraper-server替换为服务器的实际路径,并将your-api-key-here替换为您的 UseScraper API 密钥。
用法
配置完成后,您可以通过 MCP 界面使用“scrape”工具。示例用法:
{
"name": "scrape",
"arguments": {
"url": "https://example.com",
"format": "markdown"
}
}发展
对于使用自动重建的开发:
npm run watch调试
由于 MCP 服务器通过 stdio 进行通信,调试起来可能比较困难。我们推荐使用MCP Inspector ,它以包脚本的形式提供:
npm run inspector检查器将提供一个 URL 来访问浏览器中的调试工具。
Available Tools
1 toolscrapeC
Scrape content from a webpage using UseScraper API
| Name | Required | Description | Default |
|---|---|---|---|
| url | Yes | URL to scrape | |
| format | No | Format to save crawled page content. Strongly recommended to keep as markdown for optimal AI processing (default: markdown) | |
| advanced_proxy | No | Use advanced proxy to circumvent bot detection (default: false) | |
| extract_object | No | Optional object specifying data to extract |
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
With no annotations, the description fails to disclose behavioral traits beyond the API name. It does not mention rate limits, authentication, potential failures, or what happens with the proxy option. The schema provides some param hints, but the description adds no behavioral context.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is a single sentence, concise and front-loaded. It is appropriately sized for a simple tool, though it could include a bit more detail without becoming verbose.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Given the tool has 4 parameters, nested objects, and no output schema, the description is too minimal. It does not explain what is returned or how the tool behaves in edge cases, leaving significant gaps for an agent.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema coverage is 100% with each parameter described, so the baseline is 3. The description adds minimal extra meaning beyond stating the API, so it meets the baseline without enhancing understanding.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description identifies the action (scrape) and resource (webpage), naming the API used (UseScraper). However, 'scrape' is broad and could be more specific about what types of content or how it extracts, but without siblings it's sufficiently clear.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
No guidance is provided on when to use this tool versus alternatives or any exclusions. The description only states what it does, leaving the agent to infer usage context.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
Tool Schema Changelog
Recent tool additions, removals, and schema changes observed during successful MCP inspections.
1 tool update
v0.1.0- First observed
scrape
TDQS
Scored across 1 tool
Only one tool exists, so there is no possibility of confusion with other tools.
With a single tool, naming is trivially consistent.
A single tool for web scraping is borderline; it might suffice for basic use but feels thin for a full-featured scraper server.
The single scrape tool covers the core functionality, but lacks additional features like selecting specific elements or handling different output formats, which are common in web scraping APIs.
Maintenance
Related MCP Connectors
A TypeScript MCP server for Home Assistant, enabling programmatic management of entities, automati…
A simple Typescript MCP server built using the official MCP Typescript SDK and smithery/cli. This…
All HasData scraping tools in one MCP server: Google, TikTok, Instagram, maps, e-commerce and more.
One MCP server for 180+ live web-data APIs returning clean JSON from sites that block scrapers.
Related MCP Servers
- AlicenseCqualityCmaintenanceMCP Server enabling integration with Scrapezy to retrieve structured data from websites.11213MIT
- AlicenseNot gradedqualityDmaintenanceAn MCP server that enables LLMs to understand and work with TypeScript APIs they haven't been trained on by providing structured access to TypeScript type definitions and documentation.1146MIT
- AlicenseAqualityDmaintenanceA TypeScript-based web scraping server built on the Model Context Protocol that offers multiple export formats, content extraction rules, and support for both static and dynamic (SPA) websites.7121MIT
- AlicenseNot gradedqualityDmaintenanceA TypeScript-based MCP server that provides backend API handling and facilitates communication between microservices. Features an organized structure with controllers, routes, and models for easy extensibility and maintenance.2081MIT