Skip to main content
Glama

firecrawl-mcp

Server Details

Adds powerful web scraping and search to Cursor, Claude and any other LLM clients.

If you are the author of this connector, you can claim ownership by verifying the domain or GitHub account it belongs to. Claimed connector authors can inspect health checks, view analytics, and manage their listing.
Status
Healthy
Uptime
100.0% over 49 days
Last Tested
Transport
Streamable HTTP · MCP 2025-11-25
URL

TDQS

A4.2/5.0

Scored across 3 tools

Disambiguation4/5

The three tools have distinct input sources: parse for local documents, scrape for a single URL, and search for web queries. However, parse and scrape both return similar content types (markdown, HTML, links, etc.), and the verbose descriptions referencing many other tools could cause minor confusion about boundaries.

Naming Consistency5/5

All tools follow a consistent firecrawl_ prefix with a distinct verb (parse, scrape, search), yielding a predictable naming pattern throughout.

Tool Count2/5

With only 3 tools, the set is too thin for the broad web scraping, crawling, and searching scope implied by the descriptions. The descriptions reference many additional tools (crawl, map, find_tools, research_*), indicating the surface is a small subset of a much larger API.

Completeness2/5

Significant gaps exist: there are no tools for crawling multiple pages, mapping site URLs, or using the research/Alexandria capabilities mentioned in the descriptions. Agents may try to call referenced but absent tools (e.g., firecrawl_crawl, firecrawl_map) and fail.

Available Tools

3 tools
firecrawl_parseFirecrawl file parsingA
Read-only
Inspect

Parse one supported document into markdown, HTML, links, summary, targeted answers, or JSON matching a schema. Supported inputs include common HTML, PDF, Word, RTF, OpenDocument, and spreadsheet files; PDF parsing can be bounded with pdfOptions.maxPages.

Local MCP reads filePath from the server filesystem. Hosted MCP uses two calls: first provide filePath to receive upload instructions, upload locally, then call again with the returned uploadRef; do not send both fields together. Remote web URLs belong in firecrawl_scrape.

Set redactPII to request redaction of personally identifiable information in the returned content. zeroDataRetention requires an eligible authenticated account; omit it for anonymous keyless use. Returns upload instructions for hosted phase one or parsed document content for the final call. Authenticated final responses can include a data.metadata.scrapeId for optional parse feedback.

ParametersJSON Schema
NameRequiredDescriptionDefault
proxyNo
maxAgeNoIgnored: parse never reuses or stores indexed content.
formatsNo
parsersNo
filePathNoPhase 1 only: path to the local file on the caller/harness machine. Hosted MCP will not read or stat this path; it is used only to produce upload instructions.
redactPIINo
uploadRefNoPhase 2 only: short-lived upload reference returned by phase 1 after the local PUT upload completes.
pdfOptionsNo
contentTypeNoPhase 1 MIME type override. If omitted, the server infers it from the file extension without reading the file.
excludeTagsNo
includeTagsNo
jsonOptionsNo
queryOptionsNo
storeInCacheNo
onlyMainContentNo
declaredSizeBytesNoOptional phase 1 size declaration. Hosted MCP does not stat the file; provide this only if the caller already knows it.
zeroDataRetentionNo
removeBase64ImagesNo
skipTlsVerificationNo

Output Schema

ParametersJSON Schema
NameRequiredDescription
rawNoResponse body when it was not JSON.
dataNoParsed document content; can include `data.metadata.scrapeId` for parse feedback.
modeNoWhich phase of the hosted flow produced this response.
errorNoError message or error object when the call did not succeed.
notesNoHosted phase one: constraints on completing the upload flow.
uploadNoHosted phase one: how to upload the local file.
messageNoGuidance for the next call.
successNoWhether the API call succeeded.
warningNoNon-fatal warning about the result.
agent_hintsNoOptional response guidance from the Firecrawl API.
nextToolCallNoHosted phase one: the second `firecrawl_parse` call to make once the upload succeeds, as `{name, arguments}`.

TDQS

A4.5/5.0
Behavior5/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Annotations already cover the safety profile (readOnly, non-destructive, closed-world), and the description adds substantial behavior beyond them: the two-phase upload protocol, the PII redaction option, the zeroDataRetention auth requirement, and what each phase returns. This is rich context that an agent could not infer from the annotations alone.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is dense but well-structured: purpose and formats first, then the local/hosted flow distinction, then behavioral options, then return values. Every sentence carries information, with only the final scrapeId sentence being marginal.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness4/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

Given a 19-parameter tool with a nested schema and a two-phase hosted flow, the description covers the critical decision points and return behavior; an output schema exists, so return details are supplementary rather than required. Minor gaps remain around several optional tuning parameters, but nothing blocking correct invocation is missing.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is only 26% across 19 parameters, so the description must compensate. It explains filePath, uploadRef, redactPII, zeroDataRetention, pdfOptions.maxPages, and formats, but leaves proxy, parsers, contentType, includeTags/excludeTags, jsonOptions, queryOptions, storeInCache, onlyMainContent, removeBase64Images, skipTlsVerification, and declaredSizeBytes largely to the schema, which only partially documents them.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

States a specific verb (parse) and resource (one supported document) and enumerates the output formats (markdown, HTML, links, summary, query answers, schema JSON) plus supported input types. It explicitly separates itself from firecrawl_scrape by stating that remote web URLs belong there, so an agent can distinguish it from siblings without opening schemas.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines5/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

Gives explicit routing: local MCP reads filePath directly, hosted MCP requires a two-call upload flow with filePath then uploadRef, and remote URLs go to firecrawl_scrape. It also states the do-not rule (never send both fields together) and the auth condition for zeroDataRetention, which is exactly the when/when-not guidance needed.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

firecrawl_scrapeFirecrawl scrapeA
Read-only
Inspect

Scrape one URL and return its content: markdown by default, or HTML, links, screenshots, branding data, a targeted answer, or JSON matching a supplied schema. Use it when the request identifies a page and needs its content or defined fields. Use firecrawl_search when additional web sources are needed; on an authenticated session, firecrawl_map lists a site's URLs and firecrawl_crawl collects a set of pages.

Firecrawl may serve recently indexed content; set maxAge: 0 for a live fetch or a smaller maxAge to bound staleness. A successful response does not by itself confirm the page is still current. Browser actions can change the live page when interactive actions are enabled. Authenticated responses can include a metadata.scrapeId for optional scrape feedback.

On an authenticated session with Alexandria access, firecrawl_search with sources unset and firecrawl_find_tools can discover providers for the same fields across several pages; a matching provider returns typed records in one call. Keyless sessions have no provider matches.

Alexandria mode, on an authenticated session with Alexandria access: alexandria selects catalogued capability execution and is mutually exclusive with url.

ParametersJSON Schema
NameRequiredDescriptionDefault
urlNo
proxyNo
maxAgeNo
mobileNo
formatsNo
parsersNo
profileNo
timeoutNoExecution timeout in milliseconds.
waitForNo
locationNo
lockdownNo
redactPIINo
requestIdNoIdempotency key bound to one Alexandria execution payload. Generated when omitted and returned with the result.
alexandriaNoCatalogued Alexandria capability invocation, mutually exclusive with url. One {provider, capability, options} object or an array of 1-10, with contracts available through firecrawl_search or firecrawl_find_tools. Each call may include version to pin a published workflow; omitting it uses latest. Only requestId and timeout are supported alongside alexandria. The selected contract marks required inputs and any requiresOneOf groups (at least one member per group); it may include example.request/example.response and response.key (which may differ from records). Where pagination is declared, its fields govern paging with the same filters; catalogue next is separate from provider pagination. Returns per-capability results in data.alexandria with data, records, or an error with a code; individual capabilities can fail even when the outer response succeeds. Requires an API key on a team with Alexandria enabled. Some providers require accepted terms; blocked requests return the applicable requirements.
pdfOptionsNo
toolDetailNoURL mode only: domain discovery detail, summary by default; compact returns provider/capability/description, full includes contracts. Ignored with alexandria.
domainToolsNoURL mode only: include domain-matched Alexandria tools for the page in tools on the returned document. Ignored with alexandria.
excludeTagsNo
includeTagsNo
jsonOptionsNo
queryOptionsNo
storeInCacheNo
onlyMainContentNo
screenshotOptionsNo
zeroDataRetentionNo
removeBase64ImagesNo
skipTlsVerificationNo

Output Schema

ParametersJSON Schema
NameRequiredDescription
dataNoAlexandria mode: per-capability results in `data.alexandria`, each with `data`, `records`, or an `error`.
htmlNoProcessed HTML of the page.
jsonNoStructured data matching the requested JSON schema or prompt.
menuNoMenu data extracted from the page.
audioNoAudio extracted from the page.
errorNoError message or error object when the call did not succeed.
linksNoLinks found on the page.
pagesNoPhysical PDF pages, when `parsers[].pages` is set.
toolsNoDomain-matched Alexandria tools for the page, when `domainTools` is set.
videoNoVideo extracted from the page.
answerNoTargeted answer to the question that was asked of the page.
blocksNoTyped PDF layout blocks, when `parsers[].blocks` is set.
imagesNoImages found on the page.
actionsNoResults of the browser actions that ran during the scrape.
messageNoGuidance that accompanies the result.
productNoProduct data extracted from the page.
rawHtmlNoUnprocessed HTML of the page.
receiptNoBilling receipt for the execution.
successNoWhether the API call succeeded.
summaryNoSummary of the page content.
warningNoNon-fatal warning about the result.
brandingNoBranding data extracted from the page.
deliveryNo`retained` when the full result stayed server-side instead of being inlined.
markdownNoPage content as markdown.
metadataNoPage metadata; authenticated responses can include `metadata.scrapeId` for scrape feedback.
nextToolNoA follow-up tool call (`{name, arguments}`) that continues or inspects this result.
requestIdNoIdentifier of this Alexandria execution.
scrape_idNoIdentifier of the underlying scrape.
attributesNoValues collected by the requested attribute selectors.
highlightsNoHighlighted passages from the page.
screenshotNoScreenshot of the page.
agent_hintsNoOptional response guidance from the Firecrawl API.
creditsCostNoCredits this call consumed.
workspaceIdNoWorkspace holding a retained result, for inspection through virtual Bash.
feedbackToolNoPointer to the feedback tool for reporting how this result served the task.
responseBytesNoSize of the full result in bytes.
changeTrackingNoChange-tracking comparison against the previous scrape.
idleTtlSecondsNoSeconds a retained workspace stays available while idle.
estimatedTokensNoEstimated token cost of the full result.
inlineTokenBudgetNoToken budget above which a result is retained rather than inlined.
tokenEstimateMethodNoHow `estimatedTokens` was derived.

TDQS

A4.5/5.0
Behavior5/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Beyond the read-only annotations, the description discloses non-obvious behavior: Firecrawl may serve recently indexed content, maxAge controls staleness, a successful response does not confirm freshness, and browser actions can mutate the live page. It also notes the authenticated-only scrapeId and the Alexandria API-key/terms requirements.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The core capability and routing guidance are front-loaded and every paragraph carries information, but the third and fourth paragraphs (Alexandria provider discovery, catalogue paging) are dense and overlap with what the alexandria schema description already states.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness4/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a complex open-world tool with a rich output schema and strong annotations, the description covers mode selection, auth prerequisites, and freshness caveats well. The remaining gap is the large set of undocumented behavior-tuning parameters, which an agent would have to infer from names alone.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is only 19% across 27 parameters, so the description has to compensate and only partly does: it explains maxAge semantics, the formats enumeration, and the alexandria/url exclusivity and payload shape. Parameters like proxy, mobile, parsers, profile, waitFor, location, lockdown, redactPII, screenshotOptions, and jsonOptions/queryOptions remain unexplained anywhere.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The first sentence states a specific verb and resource ('Scrape one URL and return its content') and enumerates the output formats, so an agent knows exactly what comes back. It also explicitly distinguishes itself from firecrawl_search, firecrawl_map, and firecrawl_crawl, which is what separates it from siblings.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines5/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

It states the selecting condition clearly ('Use it when the request identifies a page and needs its content or defined fields') and names alternatives with their triggering conditions ('Use firecrawl_search when additional web sources are needed'), plus the map/crawl split. Alexandria mode's mutual exclusivity with url is also spelled out.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Tool Schema Changelog

Recent tool additions, removals, and schema changes observed during successful MCP inspections.

  1. 1 tool update
    • Changedfirecrawl_search2 fields changed
      • addedInput schema / properties / clientModel
        Added value: +{
        +  "description": "Optional model identifier, if known.",
        +  "maxLength": 128,
        +  "minLength": 1,
        +  "type": "string"
        +}
      • addedInput schema / properties / objective
        Added value: +{
        +  "description": "Optional broader goal for this search, if known. Avoid sensitive information.",
        +  "maxLength": 5000,
        +  "minLength": 1,
        +  "type": "string"
        +}
  2. 3 tool updates
    • Changedfirecrawl_parse1 field changed
      • addedOutput schema / properties / agent_hints
        Added value: +{
        +  "description": "Optional response guidance from the Firecrawl API.",
        +  "items": {
        +    "type": "string"
        +  },
        +  "type": "array"
        +}
    • Changedfirecrawl_scrape1 field changed
      • addedOutput schema / properties / agent_hints
        Added value: +{
        +  "description": "Optional response guidance from the Firecrawl API.",
        +  "items": {
        +    "type": "string"
        +  },
        +  "type": "array"
        +}
    • Changedfirecrawl_search1 field changed
      • addedOutput schema / properties / agent_hints
        Added value: +{
        +  "description": "Optional response guidance from the Firecrawl API.",
        +  "items": {
        +    "type": "string"
        +  },
        +  "type": "array"
        +}
  3. 1 tool update
    • Changedfirecrawl_scrape3 fields changed
      • changedInput schema / properties / alexandria / description
        Previous value: -"Catalogued Alexandria capability invocation, mutually exclusive with url. One {provider, capability, options} object or an array of 1-10, with contracts available through firecrawl_search or firecrawl_find_tools. Each call may include version to pin a published workflow; omitting it uses latest. Only requestId and timeout are supported alongside alexandria. The selected contract marks required inputs and any requiresOneOf groups (at least one member per group); it may include example.request/example.response and response.key (which may differ from records). Where pagination is declared, its fields govern paging with the same filters; catalogue next is separate from provider pagination. Returns per-capability results in data.alexandria with data, records, or an error with a code; individual capabilities can fail even when the outer response succeeds. nextTool identifies access to retained results without repeating a successful provider call. Requires an API key on a team with Alexandria enabled. THIRD_PARTY_DATA_TERMS_REQUIRED (403) means execution is blocked by provider terms, with requiresAction.url for review by an organization admin. Through this tool, terms/show displays the agreement and terms/accept records acceptance; acceptance requires explicit user authorization for the reviewed version and digest and confirmed:true. A data request does not authorize acceptance. Provider execution remains blocked until acceptance is confirmed."New value: +"Catalogued Alexandria capability invocation, mutually exclusive with url. One {provider, capability, options} object or an array of 1-10, with contracts available through firecrawl_search or firecrawl_find_tools. Each call may include version to pin a published workflow; omitting it uses latest. Only requestId and timeout are supported alongside alexandria. The selected contract marks required inputs and any requiresOneOf groups (at least one member per group); it may include example.request/example.response and response.key (which may differ from records). Where pagination is declared, its fields govern paging with the same filters; catalogue next is separate from provider pagination. Returns per-capability results in data.alexandria with data, records, or an error with a code; individual capabilities can fail even when the outer response succeeds. Requires an API key on a team with Alexandria enabled. Some providers require accepted terms; blocked requests return the applicable requirements."
      • changedInput schema / properties / requestId / description
        Previous value: -"Identifies one logical Alexandria execution; generated when omitted and returned with the result. Repeated attempts of the identical payload require the same ID. A new ID cannot reconcile a pending or uncertain execution. A caller-supplied ID supports recovery if no response is received. Errors relay a code and may include chargeId. request_in_flight (409) means the execution is pending; request_unresolved (503) requires reconciliation under the same ID; duplicate_request (409) means the ID belongs to a different payload; unknown_provider (404), insufficient_credits (402) and billing_unavailable (503) mean nothing executed."New value: +"Idempotency key bound to one Alexandria execution payload. Generated when omitted and returned with the result."
      • changedOutput schema / properties / requestId / description
        Previous value: -"Identifier of this logical execution; reuse it only for a retry of the identical payload."New value: +"Identifier of this Alexandria execution."
  4. 1 tool update
    • Changedfirecrawl_scrape2 fields changed
      • changedInput schema / properties / domainTools / description
        Previous value: -"URL mode only: include domain-matched Alexandria tools for the page in tools on the returned document."New value: +"URL mode only: include domain-matched Alexandria tools for the page in tools on the returned document. Ignored with alexandria."
      • changedInput schema / properties / toolDetail / description
        Previous value: -"URL domain discovery detail: summary by default, compact returns provider/capability/description, full includes contracts."New value: +"URL mode only: domain discovery detail, summary by default; compact returns provider/capability/description, full includes contracts. Ignored with alexandria."
  5. 2 tool updates
    • Changedfirecrawl_scrape2 fields changed
      • changedInput schema / properties / alexandria / description
        Previous value: -"Execute catalogued Alexandria capabilities instead of scraping a URL. Exactly one of url or alexandria. One {provider, capability, options} object or an array of 1-10, found through firecrawl_search or firecrawl_find_tools. Each call may include version to pin a published workflow; omitting it uses latest. Only timeout also applies at the top level. Read the selected contract before executing: required inputs and requiresOneOf groups (at least one member per group), example.request/example.response when present, and response.key (do not assume records is the result key). Follow the declared pagination input and response cursor, preserving filters; catalogue next is separate from provider pagination. Returns per-capability results in data.alexandria with data, records, or an error with a code; check each item even when the outer response succeeds. If a response provides nextTool, follow it to read a large result instead of repeating a successful provider call. Needs an API key on a team with Alexandria enabled. A terms-gated provider returns THIRD_PARTY_DATA_TERMS_REQUIRED (403) with requiresAction.url: follow the returned terms/show and terms/accept calls through this tool, accepting only after explicit user authorization for the reviewed version and digest; an organization admin can instead accept at the dashboard URL. Retry only after confirmed acceptance."New value: +"Catalogued Alexandria capability invocation, mutually exclusive with url. One {provider, capability, options} object or an array of 1-10, with contracts available through firecrawl_search or firecrawl_find_tools. Each call may include version to pin a published workflow; omitting it uses latest. Only requestId and timeout are supported alongside alexandria. The selected contract marks required inputs and any requiresOneOf groups (at least one member per group); it may include example.request/example.response and response.key (which may differ from records). Where pagination is declared, its fields govern paging with the same filters; catalogue next is separate from provider pagination. Returns per-capability results in data.alexandria with data, records, or an error with a code; individual capabilities can fail even when the outer response succeeds. nextTool identifies access to retained results without repeating a successful provider call. Requires an API key on a team with Alexandria enabled. THIRD_PARTY_DATA_TERMS_REQUIRED (403) means execution is blocked by provider terms, with requiresAction.url for review by an organization admin. Through this tool, terms/show displays the agreement and terms/accept records acceptance; acceptance requires explicit user authorization for the reviewed version and digest and confirmed:true. A data request does not authorize acceptance. Provider execution remains blocked until acceptance is confirmed."
      • changedInput schema / properties / requestId / description
        Previous value: -"Alexandria execution ID. Reuse the returned ID for retries of the identical payload, never a new ID to bypass pending or uncertain execution; generated when omitted. For potentially large workflow results, supply and preserve one before execution. Errors relay a code and chargeId: request_in_flight (409) retry the same requestId later; request_unresolved (503) keep the requestId for reconciliation, never mint a new one; duplicate_request (409) the requestId belongs to a different payload; unknown_provider (404), insufficient_credits (402) and billing_unavailable (503) mean nothing executed."New value: +"Identifies one logical Alexandria execution; generated when omitted and returned with the result. Repeated attempts of the identical payload require the same ID. A new ID cannot reconcile a pending or uncertain execution. A caller-supplied ID supports recovery if no response is received. Errors relay a code and may include chargeId. request_in_flight (409) means the execution is pending; request_unresolved (503) requires reconciliation under the same ID; duplicate_request (409) means the ID belongs to a different payload; unknown_provider (404), insufficient_credits (402) and billing_unavailable (503) mean nothing executed."
    • Changedfirecrawl_search1 field changed
      • changedInput schema / properties / sources / description
        Previous value: -"Search sources; authenticated sessions default to web + alexandria, keyless sessions to web only. Passing sources without alexandria in it (for example [\"web\"] or [\"news\"]) excludes Alexandria provider matches; omit sources unless you specifically need web-only or news-only results, or include \"alexandria\" alongside them. Use [\"alexandria\"] alone for provider discovery without web results."New value: +"Search sources; authenticated sessions default to web + alexandria, keyless sessions to web only. A search with sources: [\"web\"] omits semantic provider discovery; domainTools: true can still return website-matched tools. Web-only results use domainTools: false. Use [\"alexandria\"] alone for provider discovery without web results."
  6. 3 tool updates
    • Changedfirecrawl_parse1 field changed
      • changedOutput schema / (root)
        Previous value: -nullNew value: +{
        +  "$schema": "http://json-schema.org/draft-07/schema#",
        +  "additionalProperties": false,
        +  "description": "Parsed document content, or the upload instructions for the hosted two-call flow.",
        +  "properties": {
        +    "data": {
        +      "description": "Parsed document content; can include `data.metadata.scrapeId` for parse feedback."
        +    },
        +    "error": {
        +      "description": "Error message or error object when the call did not succeed."
        +    },
        +    "message": {
        +      "anyOf": [
        +        {
        +          "type": "string"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Guidance for the next call."
        +    },
        +    "mode": {
        +      "anyOf": [
        +        {
        +          "type": "string"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Which phase of the hosted flow produced this response."
        +    },
        +    "nextToolCall": {
        +      "description": "Hosted phase one: the second `firecrawl_parse` call to make once the upload succeeds, as `{name, arguments}`."
        +    },
        +    "notes": {
        +      "anyOf": [
        +        {
        +          "items": {
        +            "type": "string"
        +          },
        +          "type": "array"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Hosted phase one: constraints on completing the upload flow."
        +    },
        +    "raw": {
        +      "anyOf": [
        +        {
        +          "type": "string"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Response body when it was not JSON."
        +    },
        +    "success": {
        +      "anyOf": [
        +        {
        +          "type": "boolean"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Whether the API call succeeded."
        +    },
        +    "upload": {
        +      "additionalProperties": false,
        +      "description": "Hosted phase one: how to upload the local file.",
        +      "properties": {
        +        "command": {
        +          "anyOf": [
        +            {
        +              "type": "string"
        +            },
        +            {
        +              "type": "null"
        +            }
        +          ],
        +          "description": "Local command that performs the upload. It carries no Firecrawl API key."
        +        },
        +        "expiresAt": {
        +          "anyOf": [
        +            {
        +              "type": "string"
        +            },
        +            {
        +              "type": "null"
        +            }
        +          ],
        +          "description": "When the upload URL expires."
        +        },
        +        "fields": {
        +          "description": "Form fields the upload request must send, for a POST upload."
        +        },
        +        "headers": {
        +          "description": "Headers the upload request must send."
        +        },
        +        "maxSizeBytes": {
        +          "anyOf": [
        +            {
        +              "type": "number"
        +            },
        +            {
        +              "type": "null"
        +            }
        +          ],
        +          "description": "Largest file the upload URL accepts."
        +        },
        +        "method": {
        +          "anyOf": [
        +            {
        +              "type": "string"
        +            },
        +            {
        +              "type": "null"
        +            }
        +          ],
        +          "description": "HTTP method for the upload."
        +        },
        +        "uploadRef": {
        +          "anyOf": [
        +            {
        +              "type": "string"
        +            },
        +            {
        +              "type": "null"
        +            }
        +          ],
        +          "description": "Reference to pass back on the second call."
        +        },
        +        "uploadUrl": {
        +          "anyOf": [
        +            {
        +              "type": "string"
        +            },
        +            {
        +              "type": "null"
        +            }
        +          ],
        +          "description": "URL to upload the local file to."
        +        }
        +      },
        +      "type": "object"
        +    },
        +    "warning": {
        +      "anyOf": [
        +        {
        +          "type": "string"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Non-fatal warning about the result."
        +    }
        +  },
        +  "type": "object"
        +}
    • Changedfirecrawl_scrape1 field changed
      • changedOutput schema / (root)
        Previous value: -nullNew value: +{
        +  "$schema": "http://json-schema.org/draft-07/schema#",
        +  "additionalProperties": false,
        +  "description": "A scraped document, or the Alexandria execution envelope when `alexandria` was passed.",
        +  "properties": {
        +    "actions": {
        +      "description": "Results of the browser actions that ran during the scrape."
        +    },
        +    "answer": {
        +      "anyOf": [
        +        {
        +          "type": "string"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Targeted answer to the question that was asked of the page."
        +    },
        +    "attributes": {
        +      "description": "Values collected by the requested attribute selectors."
        +    },
        +    "audio": {
        +      "description": "Audio extracted from the page."
        +    },
        +    "blocks": {
        +      "description": "Typed PDF layout blocks, when `parsers[].blocks` is set."
        +    },
        +    "branding": {
        +      "description": "Branding data extracted from the page."
        +    },
        +    "changeTracking": {
        +      "description": "Change-tracking comparison against the previous scrape."
        +    },
        +    "creditsCost": {
        +      "anyOf": [
        +        {
        +          "type": "number"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Credits this call consumed."
        +    },
        +    "data": {
        +      "description": "Alexandria mode: per-capability results in `data.alexandria`, each with `data`, `records`, or an `error`."
        +    },
        +    "delivery": {
        +      "anyOf": [
        +        {
        +          "type": "string"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "`retained` when the full result stayed server-side instead of being inlined."
        +    },
        +    "error": {
        +      "description": "Error message or error object when the call did not succeed."
        +    },
        +    "estimatedTokens": {
        +      "anyOf": [
        +        {
        +          "type": "number"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Estimated token cost of the full result."
        +    },
        +    "feedbackTool": {
        +      "description": "Pointer to the feedback tool for reporting how this result served the task."
        +    },
        +    "highlights": {
        +      "description": "Highlighted passages from the page."
        +    },
        +    "html": {
        +      "anyOf": [
        +        {
        +          "type": "string"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Processed HTML of the page."
        +    },
        +    "idleTtlSeconds": {
        +      "anyOf": [
        +        {
        +          "type": "number"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Seconds a retained workspace stays available while idle."
        +    },
        +    "images": {
        +      "description": "Images found on the page."
        +    },
        +    "inlineTokenBudget": {
        +      "anyOf": [
        +        {
        +          "type": "number"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Token budget above which a result is retained rather than inlined."
        +    },
        +    "json": {
        +      "description": "Structured data matching the requested JSON schema or prompt."
        +    },
        +    "links": {
        +      "description": "Links found on the page."
        +    },
        +    "markdown": {
        +      "anyOf": [
        +        {
        +          "type": "string"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Page content as markdown."
        +    },
        +    "menu": {
        +      "description": "Menu data extracted from the page."
        +    },
        +    "message": {
        +      "anyOf": [
        +        {
        +          "type": "string"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Guidance that accompanies the result."
        +    },
        +    "metadata": {
        +      "description": "Page metadata; authenticated responses can include `metadata.scrapeId` for scrape feedback."
        +    },
        +    "nextTool": {
        +      "description": "A follow-up tool call (`{name, arguments}`) that continues or inspects this result."
        +    },
        +    "pages": {
        +      "description": "Physical PDF pages, when `parsers[].pages` is set."
        +    },
        +    "product": {
        +      "description": "Product data extracted from the page."
        +    },
        +    "rawHtml": {
        +      "anyOf": [
        +        {
        +          "type": "string"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Unprocessed HTML of the page."
        +    },
        +    "receipt": {
        +      "description": "Billing receipt for the execution."
        +    },
        +    "requestId": {
        +      "anyOf": [
        +        {
        +          "type": "string"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Identifier of this logical execution; reuse it only for a retry of the identical payload."
        +    },
        +    "responseBytes": {
        +      "anyOf": [
        +        {
        +          "type": "number"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Size of the full result in bytes."
        +    },
        +    "scrape_id": {
        +      "anyOf": [
        +        {
        +          "type": "string"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Identifier of the underlying scrape."
        +    },
        +    "screenshot": {
        +      "description": "Screenshot of the page."
        +    },
        +    "success": {
        +      "anyOf": [
        +        {
        +          "type": "boolean"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Whether the API call succeeded."
        +    },
        +    "summary": {
        +      "anyOf": [
        +        {
        +          "type": "string"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Summary of the page content."
        +    },
        +    "tokenEstimateMethod": {
        +      "anyOf": [
        +        {
        +          "type": "string"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "How `estimatedTokens` was derived."
        +    },
        +    "tools": {
        +      "description": "Domain-matched Alexandria tools for the page, when `domainTools` is set."
        +    },
        +    "video": {
        +      "description": "Video extracted from the page."
        +    },
        +    "warning": {
        +      "anyOf": [
        +        {
        +          "type": "string"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Non-fatal warning about the result."
        +    },
        +    "workspaceId": {
        +      "anyOf": [
        +        {
        +          "type": "string"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Workspace holding a retained result, for inspection through virtual Bash."
        +    }
        +  },
        +  "type": "object"
        +}
    • Changedfirecrawl_search1 field changed
      • changedOutput schema / (root)
        Previous value: -nullNew value: +{
        +  "$schema": "http://json-schema.org/draft-07/schema#",
        +  "additionalProperties": false,
        +  "description": "Ranked search results grouped by source.",
        +  "properties": {
        +    "creditsUsed": {
        +      "anyOf": [
        +        {
        +          "type": "number"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Credits this search consumed."
        +    },
        +    "data": {
        +      "description": "Ranked results grouped by source, such as `web`, `news`, `images`, and `alexandria`."
        +    },
        +    "error": {
        +      "description": "Error message or error object when the call did not succeed."
        +    },
        +    "feedbackTool": {
        +      "description": "Pointer to the feedback tool for this search."
        +    },
        +    "id": {
        +      "anyOf": [
        +        {
        +          "type": "string"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Search identifier, for optional `firecrawl_search_feedback`."
        +    },
        +    "nextTool": {
        +      "description": "A follow-up tool call that continues this search."
        +    },
        +    "success": {
        +      "anyOf": [
        +        {
        +          "type": "boolean"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Whether the API call succeeded."
        +    },
        +    "tools": {
        +      "description": "Domain-matched Alexandria tools for the results."
        +    },
        +    "warning": {
        +      "anyOf": [
        +        {
        +          "type": "string"
        +        },
        +        {
        +          "type": "null"
        +        }
        +      ],
        +      "description": "Non-fatal warning about the result."
        +    }
        +  },
        +  "type": "object"
        +}
  7. 2 tool updates
    • Changedfirecrawl_scrape2 fields changed
      • changedInput schema / properties / alexandria / description
        Previous value: -"Execute catalogued Alexandria capabilities instead of scraping a URL. Exactly one of url or alexandria."New value: +"Execute catalogued Alexandria capabilities instead of scraping a URL. Exactly one of url or alexandria. One {provider, capability, options} object or an array of 1-10, found through firecrawl_search or firecrawl_find_tools. Each call may include version to pin a published workflow; omitting it uses latest. Only timeout also applies at the top level. Read the selected contract before executing: required inputs and requiresOneOf groups (at least one member per group), example.request/example.response when present, and response.key (do not assume records is the result key). Follow the declared pagination input and response cursor, preserving filters; catalogue next is separate from provider pagination. Returns per-capability results in data.alexandria with data, records, or an error with a code; check each item even when the outer response succeeds. If a response provides nextTool, follow it to read a large result instead of repeating a successful provider call. Needs an API key on a team with Alexandria enabled. A terms-gated provider returns THIRD_PARTY_DATA_TERMS_REQUIRED (403) with requiresAction.url: follow the returned terms/show and terms/accept calls through this tool, accepting only after explicit user authorization for the reviewed version and digest; an organization admin can instead accept at the dashboard URL. Retry only after confirmed acceptance."
      • changedInput schema / properties / requestId / description
        Previous value: -"Alexandria execution ID. Reuse for retries of the identical payload; generated when omitted and returned with the result."New value: +"Alexandria execution ID. Reuse the returned ID for retries of the identical payload, never a new ID to bypass pending or uncertain execution; generated when omitted. For potentially large workflow results, supply and preserve one before execution. Errors relay a code and chargeId: request_in_flight (409) retry the same requestId later; request_unresolved (503) keep the requestId for reconciliation, never mint a new one; duplicate_request (409) the requestId belongs to a different payload; unknown_provider (404), insufficient_credits (402) and billing_unavailable (503) mean nothing executed."
    • Changedfirecrawl_search4 fields changed
      • addedInput schema / properties / excludeDomains / description
        Added value: +"Hostnames to leave out of results. Mutually exclusive with includeDomains."
      • addedInput schema / properties / includeDomains / description
        Added value: +"Hostnames to restrict results to. Mutually exclusive with excludeDomains."
      • changedInput schema / properties / query / description
        Previous value: -"Query for web and semantic tool discovery. Catalogue browsing is available through firecrawl_find_tools."New value: +"Query for web and semantic tool discovery. Operators include quoted phrases, `-term`, `site:host`, `inurl:term`, `intitle:term`, and `related:host`; the set is non-exhaustive. Catalogue browsing is available through firecrawl_find_tools."
      • addedInput schema / properties / scrapeOptions / description
        Added value: +"Attach page content for web results in the same call. These fetches ignore maxAge, so use firecrawl_scrape when you need a live fetch. scrapeOptions fetches web pages, never Alexandria provider tools."
  8. 1 tool update
    • Changedfirecrawl_search2 fields changed
      • changedInput schema / properties / query / description
        Previous value: -"Query for web and semantic tool discovery. Catalogue browsing is available on the full MCP surface."New value: +"Query for web and semantic tool discovery. Catalogue browsing is available through firecrawl_find_tools."
      • changedInput schema / properties / toolDetail / description
        Previous value: -"Compact by default. Compact returns only provider, capability and description; full includes contracts. Inspect selected compact tools on the full MCP surface using firecrawl_find_tools providers and capabilities."New value: +"Compact by default. Compact returns only provider, capability and description; full includes contracts. Inspect selected compact tools with firecrawl_find_tools providers and capabilities."
  9. 1 tool update
    • Changedfirecrawl_search3 fields changed
      • changedInput schema / properties / categories / description
        Previous value: -"Limit results to specific source types. `github` searches GitHub repositories, code, issues, and docs; `research` restricts ordinary web results to research-affiliated websites and returns page snippets, which is separate from the `firecrawl_research_*` tools that search paper abstracts and full text across biomedical (PubMed, bioRxiv, medRxiv) and arXiv literature; `pdf` searches PDF results; `developer` searches an index built for coding agents over repositories, GitHub issues, merged pull requests, repository READMEs, and curated documentation sites. `developer` returns hits in `data.web` with `category: \"developer\"`; the other categories also filter `data.web`."New value: +"Limit results to specific source types. `research` restricts ordinary web results to research-affiliated websites and returns page snippets, which is separate from the `firecrawl_research_*` tools that search paper abstracts and full text across biomedical (PubMed, bioRxiv, medRxiv) and arXiv literature; `pdf` searches PDF results; `developer` searches an index built for coding agents over public repositories, GitHub issues, merged pull requests, repository READMEs, and code documentation. `developer` returns hits in `data.web` with `category: \"developer\"`; the other categories also filter `data.web`."
      • changedInput schema / properties / categories / items / enum
        Previous value: -[
        -  "github",
        -  "research",
        -  "pdf",
        -  "developer"
        -]New value: +[
        +  "research",
        +  "pdf",
        +  "developer"
        +]
      • changedInput schema / properties / highlights / description
        Previous value: -"Return query-relevant highlights for each search result. Set to false to keep the original search snippets."New value: +"Return query-relevant page excerpts for web and news results when available (default). Highlights appear in web `description` and news `snippet`; otherwise, original snippets are returned. Set to false to keep the original search snippets."
  10. 1 tool update
    • Changedfirecrawl_search1 field changed
      • changedInput schema / properties / sources / description
        Previous value: -"Search sources; authenticated sessions default to web + alexandria, keyless sessions to web only. Use alexandria alone for semantic tool discovery."New value: +"Search sources; authenticated sessions default to web + alexandria, keyless sessions to web only. Passing sources without alexandria in it (for example [\"web\"] or [\"news\"]) excludes Alexandria provider matches; omit sources unless you specifically need web-only or news-only results, or include \"alexandria\" alongside them. Use [\"alexandria\"] alone for provider discovery without web results."

Related MCP Servers

  • A
    license
    A
    quality
    C
    maintenance
    Enables brand visibility monitoring across major AI platforms like ChatGPT, Claude, Gemini, and Perplexity. It allows users to track visibility scores, analyze competitor data, and receive actionable insights to improve AI-generated brand recommendations.
    16
    22 npm
    1
    MIT
  • A
    license
    Not graded
    quality
    B
    maintenance
    Browse IndustryLens's published competitive-intelligence reports and head-to-head competitor comparisons from any AI agent — real, source-backed data.
    MIT
Try in Browser

Glama MCP Gateway

Add one secure layer between your agents and this server.

Resources