@kazuph/mcp-fetch
MCP 페치
웹 콘텐츠를 가져오고 이미지를 처리하는 모델 컨텍스트 프로토콜 서버입니다. 이를 통해 Claude Desktop(또는 모든 MCP 클라이언트)이 웹 콘텐츠를 가져오고 이미지를 적절하게 처리할 수 있습니다.
빠른 시작(사용자용)
Claude Desktop과 함께 이 도구를 사용하려면 Claude Desktop 구성에 다음을 추가하기만 하면 됩니다( ~/Library/Application Support/Claude/claude_desktop_config.json ):
지엑스피1
필요할 때 도구의 최신 버전이 자동으로 다운로드되어 실행됩니다.
필수 설정
Claude의 접근성 활성화:
시스템 설정 열기
개인정보 보호 및 보안 > 접근성으로 이동하세요
"+" 버튼을 클릭하세요
응용 프로그램 폴더에서 Claude를 추가하세요
Claude의 토글을 켜세요
이 접근성 설정은 자동 클립보드 작업(Cmd+V)이 제대로 작동하는 데 필요합니다.
Related MCP server: OpenAI Agents MCP Server
특징
웹 콘텐츠 추출 : 웹 콘텐츠를 마크다운으로 자동 추출하고 포맷합니다.
기사 제목 추출 : 기사 제목을 추출하여 표시합니다.
이미지 처리 : 최적화된 웹 페이지의 이미지에 대한 선택적 처리(기본적으로 비활성화됨,
enableFetchImages: true로 활성화)페이지 매김 지원 : 텍스트와 이미지 모두에 대한 페이지 매김을 지원합니다.
JPEG 최적화 : 더 나은 성능을 위해 이미지를 JPEG로 자동 최적화합니다.
GIF 지원 : 애니메이션 GIF에서 첫 번째 프레임 추출
개발자를 위한
다음 섹션은 도구를 개발하거나 수정하려는 사람들을 위한 것입니다.
필수 조건
노드.js 18+
macOS(클립보드 작업용)
Claude Desktop( https://claude.ai/desktop 에서 설치)
tsx (
npm install -g tsx를 통해 설치)
설치
git clone https://github.com/kazuph/mcp-fetch.git
cd mcp-fetch
npm install
npm run build이미지 처리 사양
웹 콘텐츠의 이미지를 처리할 때 다음과 같은 최적화가 적용됩니다.
이미지는 품질 관리를 통해 JPEG 형식으로 변환됩니다.
기본적으로 최대 너비는 1200px로 제한됩니다.
기본적으로 최대 높이는 1600px로 제한됩니다.
더 나은 압축을 위한 크로마 서브샘플링(4:2:0)
더 작은 파일 크기를 위한 MozJPEG 최적화
구성
Claude Desktop이 설치되어 실행 중인지 확인하세요.
아직 설치하지 않았다면 tsx를 전역적으로 설치하세요.
npm install -g tsx
# or
pnpm add -g tsx~/Library/Application Support/Claude/claude_desktop_config.json에 있는 Claude Desktop 구성을 수정하세요.
Claude Desktop 메뉴를 통해 쉽게 찾을 수 있습니다.
클로드 데스크톱 열기
Mac 메뉴 막대에서 Claude를 클릭하세요
"설정"을 클릭하세요
"개발자"를 클릭하세요
MCP 클라이언트 구성에 다음을 추가하세요.
{
"tools": {
"fetch": {
"args": ["tsx", "/path/to/mcp-fetch/index.ts"]
}
}
}사용 가능한 매개변수
url(필수): 가져올 URLmaxLength(기본값: 20000): 반환할 콘텐츠의 최대 길이startIndex(기본값: 0): 콘텐츠의 시작 위치imageStartIndex(기본값: 0): 이미지 수집 시작 위치raw(기본값: false): 처리된 마크다운 대신 원시 콘텐츠를 반환합니다.imageMaxCount(기본값: 3): 요청당 처리할 최대 이미지 수imageMaxHeight(기본값: 4000): 병합된 이미지의 최대 높이imageMaxWidth(기본값: 1000): 병합된 이미지의 최대 너비imageQuality(기본값: 80): JPEG 품질(1-100)enableFetchImages(기본값: false): 이미지 가져오기 및 처리를 활성화합니다. CLAUDE 사용자만 해당: 이 기능을 활성화할 수 있습니다!!! CLAUDE에서 이미지를 가져오고 처리할 수 있습니다!!! 이 기능을 마음껏 사용해 보세요!!!!!ignoreRobotsTxt(기본값: false): robots.txt 제한을 무시합니다.
예시
기본 콘텐츠 가져오기(이미지 없음)
{
"url": "https://example.com"
}이미지 활성화로 가져오기
{
"url": "https://example.com",
"enableFetchImages": true,
"imageMaxCount": 3
}이미지 페이지 매기기
{
"url": "https://example.com",
"enableFetchImages": true,
"imageStartIndex": 3,
"imageMaxCount": 3
}노트
이 도구는 macOS 특정 클립보드 작업에 종속되기 때문에 macOS용으로만 설계되었습니다.
최적의 성능과 품질을 위해 Sharp를 사용하여 이미지를 처리합니다.
여러 개의 이미지가 발견되면 크기 제한을 고려하여 수직으로 병합됩니다.
애니메이션 GIF는 첫 번째 프레임을 추출하여 자동으로 처리됩니다.
변경 사항
v1.1.3
변경된 기본 동작: 기본적으로 이미지를 가져오지 않습니다(
enableFetchImages: false)enableFetchImages매개변수를 위해disableImages제거했습니다.
v1.1.0
기사 제목 추출 기능 추가
기사 제목을 포함하도록 응답 형식이 개선되었습니다.
MCP 응답 콘텐츠의 고정 유형 문제
v1.0.0
최초 출시
웹 콘텐츠 추출
이미지 처리 및 최적화
페이지 매김 지원
Available Tools
1 toolimageFetchA
画像取得に強いMCPフェッチツール。記事本文をMarkdown化し、ページ内の画像を抽出・最適化して返します。
新APIの既定(imagesを指定した場合)
画像: 取得してBASE64で返却(最大3枚を縦結合した1枚JPEG)
保存: しない(オプトイン)
クロスオリジン: 許可(CDN想定)
パラメータ(新API)
url: 取得先URL(必須)
images: true | { output, layout, maxCount, startIndex, size, originPolicy, saveDir }
output: "base64" | "file" | "both"(既定: base64)
layout: "merged" | "individual" | "both"(既定: merged)
maxCount/startIndex(既定: 3 / 0)
size: { maxWidth, maxHeight, quality }(既定: 1000/1600/80)
originPolicy: "cross-origin" | "same-origin"(既定: cross-origin)
text: { maxLength, startIndex, raw }(既定: 20000/0/false)
security: { ignoreRobotsTxt }(既定: false)
旧APIキー(enableFetchImages, returnBase64, saveImages, imageMax*, imageStartIndex 等)は後方互換のため引き続き受け付けます(非推奨)。
Examples(新API) { "url": "https://example.com", "images": true }
{ "url": "https://example.com", "images": { "output": "both", "layout": "both", "maxCount": 4 } }
Examples(旧API互換) { "url": "https://example.com", "enableFetchImages": true, "returnBase64": true, "imageMaxCount": 2 }
| Name | Required | Description | Default |
|---|---|---|---|
| url | Yes | ||
| maxLength | No | ||
| startIndex | No | ||
| imageStartIndex | No | ||
| raw | No | ||
| imageMaxCount | No | ||
| imageMaxHeight | No | ||
| imageMaxWidth | No | ||
| imageQuality | No | ||
| enableFetchImages | No | ||
| allowCrossOriginImages | No | ||
| ignoreRobotsTxt | No | ||
| saveImages | No | ||
| returnBase64 | No | ||
| images | No | ||
| text | No | ||
| security | No |
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
With no annotations, the description carries full burden. It details behavioral traits: images are fetched, base64 returned (default), max 3 images merged into one JPEG, no save unless opted in, cross-origin allowed, and old API compatibility. Absent are rate limits or auth needs, but core behavior is transparent.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is lengthy (≈250 words) but well-structured with sections for defaults, parameters, and examples. It is front-loaded with purpose but contains redundant details (e.g., repeating default values in both text and examples). Could be more concise.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Given the tool's complexity (17 parameters, nested objects, no output schema), the description provides extensive detail on new API behavior, defaults, and legacy support. Includes examples. Does not explicitly explain return format beyond base64 and Markdown, but contextually sufficient.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema coverage is 0%, so description must compensate. It lists parameters for the new API (images object with output, layout, maxCount, etc.) and mentions old keys. It covers defaults and options, though some top-level schema parameters (e.g., maxLength, startIndex) are explained only under the text object, causing slight ambiguity.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description clearly states it is an MCP fetch tool specialized for image acquisition, converting article text to Markdown and extracting/optimizing images. It specifies verb+resource (fetch and process web pages with images) and no sibling tools exist, so no differentiation needed.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
The description explains when to use (for fetching pages with images) and provides detailed parameter behavior for both new and legacy APIs. It does not explicitly exclude scenarios but offers enough context for appropriate use.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
Tool Schema Changelog
Recent tool additions, removals, and schema changes observed during successful MCP inspections.
1 tool update
v1.6.2- Changed
imageFetch4 fields changed- added
Input schema / properties / allowCrossOriginImagesAdded value: +{ + "default": true, + "type": [ + "boolean", + "string" + ] +} - added
Input schema / properties / imagesAdded value: +{ + "anyOf": [ + { + "type": "boolean" + }, + { + "additionalProperties": false, + "properties": { + "layout": { + "enum": [ + "merged", + "individual", + "both" + ], + "type": "string" + }, + "maxCount": { + "maximum": 10, + "minimum": 0, + "type": "integer" + }, + "originPolicy": { + "enum": [ + "cross-origin", + "same-origin" + ], + "type": "string" + }, + "output": { + "enum": [ + "base64", + "file", + "both" + ], + "type": "string" + }, + "saveDir": { + "type": "string" + }, + "size": { + "additionalProperties": false, + "properties": { + "maxHeight": { + "maximum": 10000, + "minimum": 100, + "type": "integer" + }, + "maxWidth": { + "maximum": 10000, + "minimum": 100, + "type": "integer" + }, + "quality": { + "maximum": 100, + "minimum": 1, + "type": "integer" + } + }, + "type": "object" + }, + "startIndex": { + "minimum": 0, + "type": "integer" + } + }, + "type": "object" + } + ] +} - added
Input schema / properties / securityAdded value: +{ + "additionalProperties": false, + "properties": { + "ignoreRobotsTxt": { + "type": "boolean" + } + }, + "type": "object" +} - added
Input schema / properties / textAdded value: +{ + "additionalProperties": false, + "properties": { + "maxLength": { + "exclusiveMinimum": 0, + "maximum": 1000000, + "type": "integer" + }, + "raw": { + "type": "boolean" + }, + "startIndex": { + "minimum": 0, + "type": "integer" + } + }, + "type": "object" +}
1 tool update
v1.0.0- First observed
imageFetch
TDQS
Scored across 1 tool
With only one tool, there is no possibility of confusion between tools. The tool's purpose is clearly defined and distinct.
The single tool name 'imageFetch' is descriptive and follows a verb_noun pattern. There is no inconsistency since only one tool exists.
While the server has only one tool, it is a complex and well-documented tool that handles a wide range of functionality appropriate for its fetch-image purpose. The count is slightly below the typical range but not insufficient.
The tool comprehensively covers fetching web pages, extracting images, and outputting them in various formats (base64, file, merged, individual). It includes parameters for text extraction and security, leaving no apparent gaps for its stated purpose.
Maintenance
Related MCP Connectors
MCP server for Firecrawl — web search, scraping, and biomedical/arXiv paper search.
Enable secure connectivity between Sentry issues and debugging data, and LLM clients, using a Model Context Protocol (MCP) server.
Model Context Protocol server for the Apideck Unified API. Connect any MCP-compatible agent framework to 100+ accounting systems, HRIS platforms, file storage providers, and more through one integration. More information https://www.apideck.com/mcp-server
A comprehensive Model Context Protocol (MCP) server that enables AI assistants to interact with yo…
Related MCP Servers
- AlicenseBqualityDmaintenanceAn educational implementation of a Model Context Protocol server that demonstrates how to build a functional MCP server for integrating with various LLM clients like Claude Desktop.1163MIT
- FlicenseBqualityDmaintenanceA Model Context Protocol server that enables Claude users to access specialized OpenAI agents (web search, file search, computer actions) and a multi-agent orchestrator through the MCP protocol.410-
- AlicenseAqualityDmaintenanceModel Context Protocol server that enables Claude Desktop (or any MCP client) to fetch web content and process images appropriately.1529 npmMIT
- AlicenseAqualityDmaintenanceA Model Context Protocol (MCP) server that enables Claude or other LLMs to fetch content from URLs, supporting HTML, JSON, text, and images with configurable request parameters.33MIT