PageProof
Server Details
PDF text by page with hashes. 0.005 USDC/Base via x402. No OCR. Check /health for readiness.
- Status
- Healthy
- Last Tested
- Transport
- Streamable HTTP
- URL
TDQS
Scored across 1 tool
Only one tool exists, so there is no possibility of confusion with other tools. Its purpose and boundaries are clearly stated.
The single tool name, pdf_extract, follows a clear verb_noun pattern. With only one tool, consistency is trivially perfect.
A single tool is insufficient for a server named PageProof, which implies a broader PDF proofing or processing domain. However, the tool itself is well-defined.
The server offers only text extraction from digital PDFs, with no support for OCR, summaries, URL fetching, or verification. This is severely incomplete for a PDF proofing service.
Available Tools
1 toolpdf_extractExtract PDF text by pageAInspect
Extract text from a digital PDF into page-numbered JSON with SHA-256 hashes. No account or API key. 0.005 USDC on Base per successful document via x402. Maximum 2 MiB and 25 pages. No OCR, summaries, URL fetching, or verification of document claims.
| Name | Required | Description | Default |
|---|---|---|---|
| pdf_base64 | Yes | Standard padded base64 of PDF bytes, maximum 2097152 bytes. No URL or data-URI prefix. |
Output Schema
| Name | Required | Description |
|---|---|---|
| pages | Yes | |
| warnings | Yes | |
| page_count | Yes | |
| provenance | Yes | |
| schema_version | Yes | |
| character_count | Yes | |
| document_sha256 | Yes |
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Adds substantial behavioral context beyond the annotations: no account or API key required, a per-document charge of 0.005 USDC on Base via x402, and hard input limits. The pricing and limits aren't derivable from the annotation flags. It does not explain the readOnlyHint=false flag (presumably the on-chain payment), leaving one small gap in an otherwise well-disclosed profile.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
Four tight sentences, front-loaded with what the tool produces; each remaining sentence carries a distinct constraint (auth, price, limits, exclusions). No filler.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
An output schema exists, so the description only needs to state the output kind, which it does. Auth, cost, size limits, and excluded capabilities are all covered for a single-parameter tool.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema coverage is 100% for the single pdf_base64 parameter, so the baseline would be 3. The description adds the 25-page ceiling, which is not expressed anywhere in the schema and meaningfully constrains what inputs are valid.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
States a specific verb and resource ('Extract text from a digital PDF') plus the exact output shape ('page-numbered JSON with SHA-256 hashes'). No siblings exist, but the scope qualification 'digital' already signals what class of input it handles.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
Explicitly carves out when this tool does NOT apply: 'No OCR, summaries, URL fetching, or verification of document claims', plus hard limits (2 MiB, 25 pages). An agent can decide against this tool for scanned PDFs or oversized documents without any further inference.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
Tool Schema Changelog
Recent tool additions, removals, and schema changes observed during successful MCP inspections.
1 tool update
- First observed
pdf_extract
Related MCP Connectors
Page extraction, product offers and feed digests. Free preview, signed prices, x402 USDC.
13 paid x402 tools, USDC/Base. Free check: GET /preview<path>=200 sample. Copy-paste pay: /llms.txt
Search, read & publish paid essays. Pay-per-read in USDC on Base (x402); wallet-only, no account.
PDF URLs to per-page text, tables as rows, Markdown, metadata and OCR for scanned pages.
Related MCP Servers
- AlicenseAqualityBmaintenanceConvert PDFs to structured JSON. Extract invoices, bank statements, contracts, and more. Pay per call via x402 USDC.58MIT
- AlicenseNot gradedqualityDmaintenanceProvides random access to PDF contents with selective page extraction, text search, outline navigation, image extraction, and page rendering capabilities. Reduces token usage by allowing targeted content extraction instead of processing entire documents.4MIT
- AlicenseAqualityCmaintenanceMCP server for document intelligence via x402 micropayments. 6 tools: document analysis, invoice extraction, screenshot data, alt text, PII detection, sentiment analysis. Pay-per-use with USDC on Base — no API keys needed.61MIT
- AlicenseAqualityBmaintenanceVerifiable document intelligence for AI agents. Extract text, tables, and structured data from PDFs and URLs. Summarize, answer questions, check claims, and translate — all with cited evidence. Store tamper-evident evidence bundles with cryptographic signatures and on-chain attestation via Base L2. Cross-document semantic search and Q&A across named collections. Pay per call with USDC22101MIT