Skip to main content
Glama
jina-ai

Jina AI Remote MCP Server

Official
by jina-ai

parallel_read_url

Extract clean content from multiple web pages simultaneously to compare information across sources or gather data from several pages at once.

Instructions

Read multiple web pages in parallel to extract clean content efficiently. For best results, provide multiple URLs that you need to extract simultaneously. This is useful for comparing content across multiple sources or gathering information from multiple pages at once.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
urlsYesArray of URL configurations to read in parallel (maximum 5 URLs for optimal performance)
timeoutNoTimeout in milliseconds for all URL reads

Schema Changelog

Changes observed during successful MCP inspections.

  1. First observed

TDQS

A3.6/5.0
Behavior2/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

No annotations are provided, so the description carries the full burden of behavioral disclosure. While it mentions 'extract clean content efficiently' and 'optimal performance' with a maximum of 5 URLs, it doesn't disclose important behavioral traits like error handling (what happens if a URL fails?), rate limits, authentication needs, or what 'clean content' specifically means. The description adds some context but leaves significant gaps for a tool that performs web requests.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is appropriately sized with three sentences that each serve a purpose: stating the core function, providing usage guidance, and explaining benefits. It's front-loaded with the main purpose first. While efficient, the third sentence could be slightly more concise by combining the two use cases.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness3/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

Given no annotations and no output schema, the description should do more to compensate. It adequately covers the purpose and basic usage but lacks details about behavioral traits, error handling, and what 'clean content' extraction entails. For a web reading tool with potential complexity around failures and content processing, this leaves important gaps despite the good schema coverage.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 100%, so the schema already fully documents both parameters (urls array with maxItems:5 and timeout with default). The description adds marginal value by reinforcing the 'multiple URLs' concept and 'optimal performance' with up to 5 URLs, but doesn't provide additional semantic meaning beyond what's in the schema descriptions.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description clearly states the tool's purpose: 'Read multiple web pages in parallel to extract clean content efficiently.' It specifies the verb ('read'), resource ('multiple web pages'), and key behavior ('in parallel'), distinguishing it from sibling tools like 'read_url' (singular) and 'parallel_search_web' (searching rather than reading/extracting).

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines4/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description provides clear context for when to use this tool: 'For best results, provide multiple URLs that you need to extract simultaneously' and 'useful for comparing content across multiple sources or gathering information from multiple pages at once.' It implies this is for parallel extraction of multiple pages, but doesn't explicitly state when NOT to use it or name alternatives like 'read_url' for single URLs.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.