CPA Search MCP Server
CPA Search MCP Server
Ingle英文| 中文说明
Un servidor Model Context Protocol (MCP) ligero y pronto para búsquedas web en tiempo real y para extracción de contenido de páginas web.
Reutiliza de manera directa tu instancia local eistente de CLIProxyAPI (CPA) (http://127.0.0.1:8317) y las cuentas autorizadas de Google Gemini, xAI Grok, Zhipu GLM para realizar investigaciones web estructuadas en tiempo real para agentes de codificación de IA (Claude Code, CurSor, Codex, Windsurf, Trae).
🌟 Características principales
️ Cerro concifiguración and concuración: Utilía direcamente tus credenciales de CPA y de Google/GLM.
ú Soporte de múltiples motores:
gemini: Búsqueda web profunda de Google y verificación de hechos (predeterminada y recomendada).glm: Búsqueda de hechos técnicos y mas (en chino mas) con Zhipu GLM.grok: Búsqueda web de xAI en tiempo real y X/Twitter.duckduckgo: Motoor de búsqueda público gratuito de respaldo sin inevidad de claves API.auto: Selección automático e inteligine de motores y conmutación por error.
🌐 Extracción de contenido web: "Incluye una herramienta busqueda emprimaria".
Compatibilidad universal**: Funciona con Claude Code, CurSor, Codex, Windsurf, Traez y cualquier cliente MCP estándar.
️ Dependencia]: Implementación puro de Node.js sobre la librería estándar de ESM. Inicio instantáneo y huellincia de memória**.
Related MCP server: DuckDuckGo MCP Server
🤐 Inestración**: Centro configuración con Cero costo adicional;** pápido**:
Configuración con mode de consulta de Mensaje.
### 1. Registro con Clarde Code
Ejecuta el siguiente comandoen tu terminación:
GXP1
Comprueba el estado:
GXP2
### 2. Configuración para CurSor / Windsurf / Traez / Codex
Añade lo siguiente a tu archio de configuración de MCP (por ejempl, `claude_desktop_config.json` o los ajestes de MCP de CurSor):
GXP3
### Tools
"**Herramientas proporcionadas**":
### 1. `cpa_search`
Busca en la web inforación, documentación y noticias técnicas.
### **Parámetros**:
**Parámetros**: `query` the `string`, required: La consult a de búsqueda o tem a de investicación".
**Parámetros**: `engine` `string`, opcional: `gemini` | `glm` | `grok` | `duckduckgo` | `auto`.
### 2. `fetch_webpage`
Obtener extrae y texto limpio desde any URL.
### **Parámetros**:
**Parámetros**: `url` the `string`, required: The web page URL (HTTP/HTTPTPS).
### ️ Variables de entor (opcional)
| Variable | Defecto | Descripción |
| :----------------- | :---------------------- | :-------------------------- |
| `CPA_BASE_URL` | `http://127.0.0.1:8312` | Dirección del proxy local de CPA |
| `CPA_API_KEY` | *(Clave local integrada)* | Clave de autentificación de CPA |
| `CPA_SEARCH_MODEL` | `gemini-3.7-flash-high` | Modelo de IA defeterminado para la búsqueda |
| `HTTPS_PROXY` | `http://127.0.0.1:7897` | Proxy de red de salida |
### ️ Pruebas
Ecuta la suite de pruebes automatizados de extremo a extremo:
GXP4
### 📄 Licencia
[Licencia MIT](https://liceffX) © 2026 fffhxAvailable Tools
2 toolscpa_searchA
使用本地已登录的 Gemini / Grok / GLM 账号进行实时全网搜索,获取最新技术动态、文档与事实总结,并附带权威来源。
| Name | Required | Description | Default |
|---|---|---|---|
| query | Yes | 搜索关键词或想要查询的最新信息 | |
| engine | No | 指定搜索引擎后端:gemini (Google 检索,默认推荐)、glm (智谱中文检索)、grok (xAI实时检索)、duckduckgo (公共搜索)、auto (默认自动智能选择) |
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Since no annotations are provided, the description carries the full burden. It discloses a key prerequisite (using locally logged-in accounts) and notes real-time retrieval with authoritative sources. However, it does not mention potential failures (e.g., no logged-in accounts), rate limits, or result formatting, leaving gaps.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is a single, concise sentence that front-loads the core function (real-time web search) and adds relevant context (accounts, sources) without extraneous words. It is optimally sized.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Despite lacking an output schema, the description hints at return content (tech trends, documents, fact summaries, sources) and the prerequisite of logged-in accounts. It does not explain error handling or pagination, but these are less critical for a simple search tool. Overall, it provides sufficient context for an agent to call it correctly.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema description coverage is 100%, and the description does not add significant meaning beyond the schema. It mentions using logged-in accounts, which links to the 'engine' parameter, but the schema already explains each backend option. The description adds minimal value for parameters.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description explicitly states the tool performs real-time web-wide search ('实时全网搜索'), retrieves latest tech trends, documents, and fact summaries with authoritative sources. This clearly distinguishes it from the sibling 'fetch_webpage', which fetches a specific page.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
The description implies usage for searching rather than fetching specific pages, but it does not explicitly state when to use this tool over 'fetch_webpage' or provide exclusion criteria. The context of 'search' vs 'fetch' is implicit but not articulated.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
fetch_webpageC
直接抓取指定 URL 网页的完整正文内容并转换为易读的文本
| Name | Required | Description | Default |
|---|---|---|---|
| url | Yes | 需要抓取的完整网页网址(例如 https://docs.example.com) |
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
No annotations are provided, so the description must fully convey behavior. It mentions fetching the 'complete main content' and converting to text, but omits any mention of handling dynamic content, authentication, timeouts, or error cases. This is a significant gap for a web-fetching tool.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is a single sentence with no unnecessary words. It is concise and gets to the point, though it could slightly benefit from structuring key constraints more explicitly.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
For a tool with a single parameter and no output schema, the description should cover potential failure modes and limitations. It lacks any mention of what happens on failure, unsupported URL types, or rate limits. Given the lack of annotations, this is incomplete for safe agent use.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
The schema description coverage is 100%, so the url parameter is already well-documented. The description does not add extra semantic meaning beyond what the schema provides, matching the baseline for high schema coverage.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description clearly states the verb 'fetch' and the resource 'webpage', and adds that it converts content to readable text. It is specific enough, though it doesn't explicitly contrast with sibling tool cpa_search, which appears to be a different operation.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
No guidance is given on when to use this tool versus alternatives, nor any prerequisites or exclusions. The description simply states what it does, leaving the agent to infer when it is appropriate.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
Tool Schema Changelog
Recent tool additions, removals, and schema changes observed during successful MCP inspections.
2 tool updates
v1.0.0- First observed
cpa_search - First observed
fetch_webpage
TDQS
Scored across 2 tools
两个工具功能明确分离:cpa_search用于搜索并返回结果,fetch_webpage用于抓取特定URL的内容。没有重叠或模糊边界,代理可以轻松区分何时使用哪个工具。
两个工具均使用snake_case,且都包含动作词(search, fetch),但cpa_search以名词cpa开头,而fetch_webpage以动词开头,存在轻微不一致。不过整体模式简单且可预测,不影响理解。
仅2个工具对于搜索服务器而言显得单薄,但结合其功能(搜索+网页抓取)已覆盖核心需求。未超出合理范围,但处于评分的下限。
搜索和网页抓取共同构成完成查询-获取详情的基本流程,没有明显缺失的关键操作。但缺少如历史记录、批量处理等扩展功能,仍有小幅度提升空间。
Maintenance
Related MCP Connectors
LLM-ready web search + instant answers + URL-to-clean-text fetch for agents and RAG.
Web search, browser automation, scraping, crawling and CAPTCHA solving for AI agents.
Web search for AI agents — one tool across 6 engines, routed to the cheapest + cached.
Fetch pages as markdown, search web and news, extract structured data. For AI agents.
Related MCP Servers
- AlicenseNot gradedqualityDmaintenanceEnables comprehensive web searching and content extraction using multiple search engines (Bing, Brave, DuckDuckGo) without API keys. Provides tools for full web searches with content extraction, quick search summaries, and single webpage content retrieval.1,168MIT
- FlicenseNot gradedqualityCmaintenanceEnables AI agents to search the web via DuckDuckGo and fetch relevant webpage content using an LLM, without requiring an API key.-
- FlicenseNot gradedqualityDmaintenanceEnables AI agents to perform web searches, extract webpage content, and conduct end-to-end search-and-extract operations using multiple search providers and content extraction methods.-
- AlicenseNot gradedqualityCmaintenanceProvides AI coding agents with a live web search tool that returns extracted, answer-ready text from web pages. Enables real-time information retrieval without any local installation or maintenance.4 npmMIT