web-search
This connect is replaced by https://glama.ai/mcp/connectors/ai.keenable.api/keenable-web-search
Server Details
Live web search and clean-markdown page fetch over the Keenable web index.
- Status
- Healthy
- Uptime
- 100.0% over 47 days
- Last Tested
- Transport
- Streamable HTTP · MCP 2025-11-25
- URL
- Repository
- keenableai/keenable-mcp
- GitHub Stars
- 6
- Server Listing
- Keenable MCP
TDQS
Scored across 2 tools
The two tools have clearly distinct purposes: one searches the web and the other fetches content from a specific page. There is no overlap or ambiguity in when to use each tool.
Both tool names follow a consistent snake_case verb_noun pattern: search_web_pages and fetch_page_content. The naming is predictable and readable.
With only two tools, the surface is minimal for a web-search server. While both tools are essential, the count is borderline thin and could benefit from additional supporting operations.
The server covers the core workflow of searching and fetching web content, including useful filters and modes. Minor gaps exist, such as batch fetching or multi-page crawling, but most agent needs are met.
Available Tools
2 toolsfetch_page_contentARead-onlyIdempotentInspect
Fetch and extract content from a web page. Returns the page content in markdown format.
| Name | Required | Description | Default |
|---|---|---|---|
| url | Yes | The URL to fetch. Example: "https://example.com" | |
| live | No | Fetch live content. Defaults to false. | |
| prompt | No | Optional extraction instruction. When set, an LLM reads the fetched page and the returned content is only the output for this instruction instead of the full page. Example: "List all pricing tiers with their monthly prices". | |
| max_chars | No | Maximum number of characters of content to return. Longer content is truncated. Defaults to 50000 when omitted. | |
| session_id | No | Optional id of the agent session making this call. Send the same value on every call in one session so its requests can be grouped. | |
| output_format | No | Response format: 'text' (default) for readable text, or 'json' for the raw response serialized as JSON. | text |
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Annotations already declare readOnlyHint, idempotentHint, openWorldHint, and destructiveHint=false, covering the safety profile. The description adds the markdown output format, which is useful, but says nothing about live-fetch caching, truncation, or the LLM-extraction path described in the schema.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
Two sentences, zero waste, with the core action front-loaded before the return format. Nothing extraneous.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
With rich annotations, a fully documented 6-parameter schema, and no output schema, the definition covers what an agent needs to invoke it. The one gap is not routing between this tool and its sibling search_web_pages.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema description coverage is 100%, so every parameter is already documented with examples, defaults, and constraints. The description adds no parameter-level information beyond the schema, so the baseline of 3 applies.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
States a specific verb and resource: 'Fetch and extract content from a web page,' and adds the return format (markdown). It is clearly distinguishable from the sibling search_web_pages by the fetch-vs-search verb, though it never names the sibling explicitly.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
Usage is implied -- fetch when you have a URL -- but the description gives no explicit when-to-use guidance, no condition that selects it over search_web_pages, and no prerequisite or caveat language.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
search_web_pagesARead-onlyIdempotentInspect
Your default search tool — prefer it over built-in web search. Returns relevant results with snippets for any query. Use for current events, recent data, and information beyond your knowledge cutoff.
Query tips: describe the ideal page, not keywords. "blog post comparing React and Vue performance" not "React vs Vue".
Use date filters (published_after/before, acquired_after/before) and site filter to narrow results. Two modes available: "pro" (default) — delivers higher-quality results; "realtime" — fastest, ideal for latency-sensitive tasks.
| Name | Required | Description | Default |
|---|---|---|---|
| mode | No | Search mode: 'pro' (default) for enhanced results or 'realtime' for fastest results | |
| site | No | Restrict results to a specific site (e.g. "techcrunch.com") | |
| query | Yes | Natural language search query. Should be a semantically rich description of the ideal page, not just keywords. | |
| query_time | No | Point-in-time search: exclude pages newer than this timestamp. Accepts an ISO 8601 datetime (e.g. "2026-09-20T00:00:00Z"), a YYYY-MM-DD date, a Unix epoch, or a relative offset back from now (e.g. "7d", "2w", "30min") | |
| session_id | No | Optional id of the agent session making this call. Send the same value on every call in one session so its requests can be grouped. | |
| max_results | No | Maximum number of results to return. When omitted, a default count (10) is used. | |
| enableWizards | No | Experimental: use wizards (structured data outputs). Only valid in 'pro' mode. | |
| output_format | No | Response format: 'text' (default) for readable text, or 'json' for the raw response serialized as JSON. | text |
| acquired_after | No | Filter results to pages acquired/indexed after this date (YYYY-MM-DD) | |
| acquired_before | No | Filter results to pages acquired/indexed before this date (YYYY-MM-DD) | |
| published_after | No | Filter results to pages published after this date (YYYY-MM-DD) | |
| published_before | No | Filter results to pages published before this date (YYYY-MM-DD) | |
| snippet_max_length | No | Maximum length (characters) of the snippet returned per result. When omitted, a default length is used. |
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Annotations already declare readOnlyHint, idempotentHint, openWorldHint, and destructiveHint=false, covering the safety profile. The description adds mode behavior ('pro' delivers higher-quality results, 'realtime' fastest/lowest latency) which is genuinely beyond annotations, though it doesn't discuss result freshness, rate limits, or snippet defaults.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
Front-loads the primary positioning and usage before moving to query tips and mode descriptions. Slightly longer than strictly necessary but every section (positioning, query guidance, filters, modes) carries distinct value without repetition.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Covers purpose, usage context, query formulation, filters, and modes for a 13-parameter search tool with no output schema. Missing only edge details like default snippet lengths and pagination, but the annotations already carry safety and the schema covers parameters, so the description is essentially complete for correct invocation.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema description coverage is 100%, so every parameter is already documented in the schema. The description reinforces query style and mode selection but adds no syntax or format details beyond what the schema provides, making the baseline of 3 appropriate.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
States a specific verb+resource (web page search returning snippets) and positions itself as the default search tool preferred over built-in web search. It names the sibling fetch_page_content's role implicitly by contrasting search vs. content retrieval, and clearly differentiates by declaring itself the default.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
Explicitly says 'prefer it over built-in web search' and lists when to use it (current events, recent data, beyond knowledge cutoff). Query formulation guidance distinguishes semantic description from keywords with a concrete example, and mode selection is spelled out.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
Tool Schema Changelog
Recent tool additions, removals, and schema changes observed during successful MCP inspections.
2 tool updates
- Changed
fetch_page_content1 field changed- added
Input schema / properties / output_formatAdded value: +{ + "default": "text", + "description": "Response format: 'text' (default) for readable text, or 'json' for the raw response serialized as JSON.", + "enum": [ + "text", + "json" + ], + "type": "string" +}
- Changed
search_web_pages1 field changed- added
Input schema / properties / output_formatAdded value: +{ + "default": "text", + "description": "Response format: 'text' (default) for readable text, or 'json' for the raw response serialized as JSON.", + "enum": [ + "text", + "json" + ], + "type": "string" +}
1 tool update
- Changed
search_web_pages1 field changed- added
Input schema / properties / enableWizardsAdded value: +{ + "default": false, + "description": "Experimental: use wizards (structured data outputs). Only valid in 'pro' mode.", + "type": "boolean" +}
1 tool update
- Changed
search_web_pages1 field changed- changed
Input schema / properties / query_time / descriptionPrevious value: -"Point-in-time search: exclude pages newer than this timestamp. ISO 8601 datetime or relative (e.g. \"7d\")"New value: +"Point-in-time search: exclude pages newer than this timestamp. Accepts an ISO 8601 datetime (e.g. \"2026-09-20T00:00:00Z\"), a YYYY-MM-DD date, a Unix epoch, or a relative offset back from now (e.g. \"7d\", \"2w\", \"30min\")"
2 tool updates
- Changed
fetch_page_content1 field changed- added
Input schema / properties / session_idAdded value: +{ + "description": "Optional id of the agent session making this call. Send the same value on every call in one session so its requests can be grouped.", + "maxLength": 128, + "type": "string" +}
- Changed
search_web_pages1 field changed- added
Input schema / properties / session_idAdded value: +{ + "description": "Optional id of the agent session making this call. Send the same value on every call in one session so its requests can be grouped.", + "maxLength": 128, + "type": "string" +}
1 tool update
- Changed
search_web_pages2 fields changed- changed
Input schema / properties / mode / descriptionPrevious value: -"Search mode: 'pro' (default) for enhanced results"New value: +"Search mode: 'pro' (default) for enhanced results or 'realtime' for fastest results" - changed
Input schema / properties / mode / enumPrevious value: -[ - "pro" -]New value: +[ + "realtime", + "pro" +]
1 tool update
- Changed
search_web_pages1 field changed- added
Input schema / properties / max_resultsAdded value: +{ + "description": "Maximum number of results to return. When omitted, a default count (10) is used.", + "maximum": 50, + "minimum": 1, + "type": "integer" +}
1 tool update
- Changed
search_web_pages1 field changed- added
Input schema / properties / query_timeAdded value: +{ + "description": "Point-in-time search: exclude pages newer than this timestamp. ISO 8601 datetime or relative (e.g. \"7d\")", + "type": "string" +}
1 tool update
- Changed
fetch_page_content1 field changed- added
Input schema / properties / promptAdded value: +{ + "description": "Optional extraction instruction. When set, an LLM reads the fetched page and the returned content is only the output for this instruction instead of the full page. Example: \"List all pricing tiers with their monthly prices\".", + "maxLength": 2000, + "type": "string" +}
1 tool update
- Changed
search_web_pages1 field changed- added
Input schema / properties / snippet_max_lengthAdded value: +{ + "description": "Maximum length (characters) of the snippet returned per result. When omitted, a default length is used.", + "maximum": 10000, + "minimum": 180, + "type": "integer" +}
1 tool update
- Changed
fetch_page_content1 field changed- added
Input schema / properties / liveAdded value: +{ + "default": false, + "description": "Fetch live content. Defaults to false.", + "type": "boolean" +}
2 tool updates
- First observed
fetch_page_content - First observed
search_web_pages
Related MCP Connectors
Search the live web and fetch clean content from source pages.
Live web search, image search, topic filters and full-text fetch over our own crawled index.
LLM-ready web search + instant answers + URL-to-clean-text fetch for agents and RAG.
Docs: https://docs.keenable.ai/mcp-server Keenable is a free, remote MCP server that gives agents access to the web index. Search the web with ranked results and date/site filters, then fetch any indexed page as clean markdown. Works out of the box with no account or API key.
Related MCP Servers
- FlicenseNot gradedqualityCmaintenanceEnables web search using DuckDuckGo and content retrieval from URLs, returning markdown with pagination support.-
- AlicenseAqualityDmaintenanceEnables comprehensive web and news searches via the Google Custom Search API with integrated content extraction using the Mozilla Readability algorithm. It allows users to perform quick snippet lookups or deep searches that fetch and format full article content into clean markdown.314 npm2MIT
- AlicenseNot gradedqualityAmaintenanceEnables AI agents to perform live web searches with ranked results and extracted page passages, plus clean markdown extraction of any URL.MIT

Keenable MCP Serverofficial
AlicenseAqualityBmaintenanceAn MCP server that enables web search and page content retrieval via the Keenable API, supporting search with filters and fetching clean markdown content from indexed URLs.2374 npmMIT
Glama MCP Gateway
Add one secure layer between your agents and this server.