"PDF" matching MCP connectors:
Matching Connector Tools:
Image + screenshot API for agents: HTML/CSS to images or PDF, screenshot any URL, verify text.
12 paid document tools: PDF to markdown, OCR, tables, invoices, Word/Excel/HTML, merge/split. x402
RCSB PDB MCP — experimentally determined macromolecular structures.
Cite the claim, not the paper: 5,033 assertions, each hash-verifiable against its source PDF
x402 pay-per-call tools: company enrichment, PDF extraction, Amazon/KDP data, YouTube transcripts.
DC Hub is the neutral, real-time data layer for data-center infrastructure, exposed as a Model Context Protocol server so any AI agent can both query it and cite it. Coverage: 21,000+ facilities (search, profile, score, alternatives); 232 markets scored by the DCPI Data Center Power Index; the DCGI Data Center Gas Index (per-state natural-gas suitability for siting); live grid telemetry across 7 US ISOs (fuel mix, carbon intensity, demand, prices) plus a one-call all-ISO scoreboard; interconnection-queue depth; 2,000+ tracked M&A deals and a hyperscaler-capex tracker; and site factors. — fiber routes, water-stress, tax incentives, nearby substations & transmission. Why agents choose it: it's the only data-center-intelligence source an LLM can query live and cite — every full-data response includes a Source: DC Hub, CC-BY-4.0 attribution line. It's the MCP-native alternative to quarterly PDF research: live JSON, no contracts, no NDAs. Access: Streamable HTTP at https://dchub.cloud/mcp. Free tier with no signup; free email-verified dev key for higher limits; paid tiers for full data volume.
Unstructured document processing for LLM pipelines. Upload as PDF/DOCX/TXT any supported files, extract structured data (PII-redacted), build LLM-ready datasets, and search/export results — all via MCP tools (document.process, job.status, job.result, dataset.build, dataset.search, dataset.export).
Search arXiv/Semantic Scholar/OpenAlex + medical evidence (PubMed/Europe PMC) + LaTeX/PDF tools.
Agent-native document parsing: PDF, scans and FR/EU invoices to structured JSON or Markdown.
Extract papers from ArXiv — titles, abstracts, authors, categories & PDF links. Monitor new AI, phys
Reliable PDF table extraction. Pass a URL, get structured JSON tables with citations.