Skip to main content
Glama
sathvic-kollu

Techtenstein PDF MCP

Techtenstein PDF MCP

MCP server that gives your Claude, Cline, or Cursor session the ability to extract text, tables, and metadata from any PDF URL — including scanned PDFs via OCR. Powered by the Techtenstein PDF Extract API.

Tools exposed

  • pdf_extract_text(pdf_url, ocr=False) — Extract all text from a PDF as clean plain text

  • pdf_extract_tables(pdf_url) — Extract all tables as structured row arrays

  • pdf_metadata(pdf_url) — Get title, author, page count, creation date, encryption status

Related MCP server: pdf-mcp

Install (Claude Desktop)

Add to ~/Library/Application Support/Claude/claude_desktop_config.json:

{
  "mcpServers": {
    "techtenstein-pdf": {
      "command": "uvx",
      "args": ["techtenstein-pdf-mcp"],
      "env": {
        "TECHTENSTEIN_API_KEY": "your_key_from_techtenstein.com"
      }
    }
  }
}

Restart Claude Desktop. pdf_extract_text, pdf_extract_tables, and pdf_metadata will appear as available tools.

Install (Cline / VS Code)

Cline auto-detects MCP servers from your Claude Desktop config. Same setup as above works.

Install (Cursor)

Add to ~/.cursor/mcp.json:

{
  "mcpServers": {
    "techtenstein-pdf": {
      "command": "uvx",
      "args": ["techtenstein-pdf-mcp"],
      "env": {"TECHTENSTEIN_API_KEY": "your_key"}
    }
  }
}

Get an API key

Free tier (50 extractions/day, no card): https://apis.techtenstein.com

Paid tiers start at $5/month for 2,000 extractions.

Example usage

Once installed, ask Claude:

"Extract the tables from this earnings report PDF: https://example.com/q4.pdf"

Claude will call pdf_extract_tables and return a clean structured view of every table on the page.

Or for scanned documents:

"This PDF is a scanned invoice. Extract the text: https://example.com/invoice.pdf"

Claude will call pdf_extract_text(pdf_url, ocr=True) and read the image-based text via OCR.

Response schema (text mode)

{
  "text": "Full extracted body text...",
  "page_count": 12,
  "word_count": 3450,
  "ms": 240
}

Response schema (tables mode)

{
  "tables": [
    {
      "page": 3,
      "rows": [
        ["Product", "Q1", "Q2", "Q3", "Q4"],
        ["Widget A", "1200", "1350", "1420", "1600"]
      ]
    }
  ]
}

Support

License

MIT

A
license - permissive license
Not graded
quality - not tested
C
maintenance

Maintenance

Maintainers
Response time
Release cycle
Releases (12mo)
Commit activity

Resources

Unclaimed servers have limited discoverability.

Looking for Admin?

If you are the server author, to access and configure the admin panel.

Related MCP Servers

  • A
    license
    A
    quality
    C
    maintenance
    An MCP server for reading, rendering, and searching PDF files, specifically optimized for LLMs to extract text, tables, and technical diagrams. It enables metadata retrieval, multi-format text extraction, and page-to-image rendering using PyMuPDF.
    5
    66
    MIT
  • A
    license
    Not graded
    quality
    A
    maintenance
    MCP server that reads PDFs and exposes them as structured Markdown, metadata, outlines, images, and tables to LLM consumers via tools like pdf_read_markdown and pdf_info.
    Apache 2.0
  • F
    license
    Not graded
    quality
    D
    maintenance
    A local MCP server that extracts text-layer content from PDF files, enabling AI agents to inspect, extract text, outlines, and page content.

View all related MCP servers

Related MCP Connectors

  • Generate PDFs from templates via AI chat. Works with Claude, ChatGPT, Cursor, and any MCP client.

  • OCR, transcription, file extraction, and image generation for AI agents via MCP.

  • Augments MCP Server - A comprehensive framework documentation provider for Claude Code

View all MCP Connectors

Latest Blog Posts

MCP directory API

We provide all the information about MCP servers via our MCP API.

curl -X GET 'https://glama.ai/api/mcp/v1/servers/sathvic-kollu/techtenstein-pdf-mcp'

If you have feedback or need assistance with the MCP directory API, please join our Discord server