Skip to main content
Glama
525,485 tools. Updated 2026-09-06 21:34

"A tool or method for searching PDF documents" matching MCP tools:

  • Converts documents between 100+ formats (Office, PDF, images, HTML) while preserving Carbone template tags. Supports PDF options like watermarks, passwords, or page ranges.
    Apache 2.0
  • Add documents to a collection by providing a URL for download, processing them for text extraction, and indexing them for semantic search.
    MIT
  • List documents in a Needle collection to check processing status, inventory available files, and verify document availability before searching.
    MIT
  • Create a new document collection to organize related documents and enable semantic search across their contents. Returns a collection ID for managing documents.
    MIT
  • Retrieve AI-generated answers by searching your namespace of text documents or using direct AI model calls for question answering.
    Apache 2.0
  • Extract figures, tables, and equations from PDF documents using layout detection. Returns base64-encoded images of detected elements with metadata from academic papers or any PDF URL.
    Apache 2.0

Matching MCP Servers

Matching MCP Connectors

  • List all available documents in the configured directory with format detection and metadata including filename, pages, and size. Discover documents before reading or searching.
    MIT
    Destructive
  • Quickly estimate the number of documents in a collection in Astra DB using a fast, approximate counting method to support efficient data management.
    Apache 2.0
  • Split one PDF into multiple documents by specifying page ranges; each range produces a separate PDF file or embedded resource.
    MIT
  • Parse local documents (PDF, Word, Excel, HTML) into markdown or structured JSON. Extract content with format selection and PDF-specific options.
    MIT
  • Perform precise searches across Dutch parliamentary records, including documents, debates, and member information. Filter results using keywords, exact phrases, or advanced operators like 'NOT', 'OR', and 'NEAR()' for accurate topic tracking.
    MIT
  • Refine search queries within Dutch parliamentary data by filtering results to specific content types such as documents, activities, or cases. Use advanced syntax for exact phrases, exclusions, alternatives, or proximity searches to obtain precise results.
    MIT
  • Extract text or images from PDF files for vision LLMs, handling both text and scanned documents with configurable extraction modes.
    MIT
  • Retrieve specific financial documents by ID to access content, metadata, and extracted PDF text for comprehensive market analysis.
    MIT
  • Find documents by topic using semantic search over summaries. Returns titles, headings, and relevance scores to help locate relevant documents before searching for specific content.
    MIT
  • Convert DOCX files to PDF, prioritizing LibreOffice for high fidelity, then pandoc. Returns conversion status and method used.
    MIT
  • Queue high-fidelity background ingest of EPUB and PDF documents. Auto-detect book or paper profile, force re-ingest if needed.
    MIT
  • Ingest documents (PDF, DOCX, TXT, MD) into a vector database for semantic search. Supports updating existing documents and visual captioning for PDF figures.
    MIT
  • Convert PDF form fields and annotations into static content to make them non-editable, finalizing documents for distribution.
    MIT
  • Embed one or more files into a PDF as attachments that travel with the document and stay extractable by readers. Bundle source data, receipts, or supporting documents directly with your PDF.
    MIT