Skip to main content
Glama
massanaRoger

extracto-mcp

by massanaRoger

Server Configuration

Describes the environment variables required to run the server.

NameRequiredDescriptionDefault
EXTRACTO_API_KEYYesYour key from app.getextracto.dev/keys.
EXTRACTO_BASE_URLNoOverride the API host (defaults to https://app.getextracto.dev).
EXTRACTO_TIMEOUT_MSNoPer-request timeout in ms (default 90000).

Instructions

Guidance the server publishes about itself, which clients place ahead of the tool catalog so the model reads it before choosing anything.

This server publishes no instructions, or was last inspected before Glama recorded them.

Capabilities

Features and capabilities supported by this server

Protocol revision2025-11-25

CapabilityDetails
tools
{
  "listChanged": true
}

Tools

Functions exposed to the LLM to take actions

NameDescription
extractA

Extract structured data from a public web page and return it as validated, typed JSON. Extracto renders the page (JavaScript included), runs a schema-constrained extraction, and returns ONLY fields that match the schema. Missing data comes back as null rather than a hallucinated guess. Best for a single known URL. This call is synchronous (up to ~90s); for heavy or anti-bot pages prefer extract_async.

extract_asyncA

Submit an asynchronous extraction job for a heavy, slow, or anti-bot-protected page. Returns a job id immediately; poll it with get_job until status is "success" or "failed". Use this instead of extract when a page is large or likely to need stealth rendering.

get_jobA

Fetch the current status and (once complete) the result of an async extraction job created with extract_async. Status is one of pending, processing, success, failed.

list_jobsA

List your recent async extraction jobs, newest first.

Prompts

Interactive templates invoked by user choice

NameDescription

No prompts

Resources

Contextual data attached and managed by the client

NameDescription

No resources

TDQS

A4.4/5.0

Scored across 4 tools

Disambiguation5/5

Each tool serves a distinct purpose: synchronous extraction, async submission, polling, and listing. No overlap or ambiguity.

Naming Consistency4/5

Names are mostly lowercase and descriptive, but 'extract_async' breaks the verb_noun pattern seen in 'get_job' and 'list_jobs'. Minor inconsistency.

Tool Count5/5

With 4 tools covering sync/async extraction and job management, the count is well-scoped for a focused extraction service.

Completeness4/5

Covers core extraction workflows, but lacks a cancel job or delete job tool. Minor gap for a complete surface.

Maintenance

ActivityInactive
ResponsivenessNo issues