Skip to main content
Glama

document-intelligence-server - MCP Server

MCP server for intelligent document processing — extract text, classify, and summarize PDFs, images, and documents using vision LLMs.

Tools

extract_document

Extract text from a PDF or image file using OCR.

  • path (string, required): Path to the document file

classify_document

Classify a document by type (invoice, report, contract, etc.).

  • path (string, required): Path to the document file

summarize_document

Generate a structured summary from a document.

  • path (string, required): Path to the document file

Related MCP server: Nutrient Document Engine MCP Server

Quick Start (local)

pip install -r requirements.txt
export MCP_BILLING_API=https://mcp-billing-api.onrender.com
uvicorn server:starlette_app --host 0.0.0.0 --port 8000

Usage with Claude Desktop / MCP clients

{
  "mcpServers": {
    "document-intelligence-server": {
      "url": "https://mcp-doc-intel.onrender.com/"
    }
  }
}

Deployed endpoint

https://mcp-doc-intel.onrender.com/ — Streamable HTTP transport at root path. Health check at /health.

Environment Variables

Variable

Required

Description

MCP_BILLING_API

Yes

Billing API endpoint

MCP_LICENSE_KEY

Yes

License key for billing

AGENTICMARKET_SECRET

No

Secret for AgenticMarket authentication

License

MIT

PyPI GitHub

Related MCP Connectors

Related MCP Servers

  • A
    license
    Not graded
    quality
    B
    maintenance
    Enables AI agents and users to process documents through natural language, supporting PDF operations like text extraction, redaction, splitting, form filling, annotations, and content search.
    42 npm
    62
    MIT
  • A
    license
    Not graded
    quality
    F
    maintenance
    Enables PDF document processing including text, image, and table extraction, as well as intelligent classification and similarity analysis across multiple languages.
    49
    MIT
  • A
    license
    Not graded
    quality
    C
    maintenance
    Provides AI agents with comprehensive document parsing capabilities including PDF text extraction, OCR, HTML-to-markdown conversion, table extraction, and summarization, optimized for agent workflows.
    30 npm
    MIT