web_links
Inspect a known HTML page to list its explicit links, helping you navigate documentation chapters, pagination, references, or related pages after search and before reading.
Instructions
Navigate from a known HTML page to pages that it explicitly links.
USE THIS when you already found a useful source but need its structure:
documentation sections, report chapters, table-of-contents entries,
next/previous publication pages, appendices, references, datasets, standards,
or related pages. It is the bridge between web_search discovery and web_fetch reading.
Prefer this over another search when the desired page is plausibly
linked from the source you already have.
This is NOT a crawler or browser. It inspects ordinary <a href> links in this
one fetched HTML page only. It does not follow them, click controls, execute
JavaScript, submit forms, DNS-resolve destinations, or safety-approve them.
Selecting a returned URL for web_fetch/web_links later triggers the normal
outbound SSRF policy.
same_origin=true keeps only links with the same scheme, normalized hostname
and effective port as the final fetched page. Leave same_origin=null when
external citations or primary sources may matter; same origin does not mean
same organization, and cross-origin does not mean untrusted.
The retained list preserves document order after deterministic URL
deduplication. Fragments remain part of link identity. same_document=true
means only that the destination has the same document identity ignoring the
fragment; web_fetch still navigates text by Unicode offsets, not HTML anchors.
Filtering happens BEFORE pagination. Copy continuation.arguments to retrieve
the next retained link page. links_hash versions the whole retained ordered
list before filtering/pagination; links_changed means the old start_index must
not be reused. continuation means more retained links exist; links_truncated
means some admissible links were lost to hard page-level resource ceilings
and continuation cannot recover them.
Each link returns url, best-effort label, rel, fragment, same_origin and
same_document. An empty links list does not prove that the site has no other
pages. Expected failures return {error, hint}.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| url | Yes | Known HTTP(S) HTML page whose exposed navigation you want to inspect. Use after finding a useful source when you need chapters, pagination, references or related pages. The page itself is fetched safely; returned destinations are not contacted until selected later. | |
| max_links | No | Maximum links returned: default 50, hard cap 100. Ready continuation preserves the effective budget. | |
| same_origin | No | Navigation filter relative to the final fetched page. true=same scheme+normalized host+effective port only; false=cross-origin only; null=keep both (default, best when external references may matter). Filtering happens before pagination. | |
| start_index | No | Zero-based index in the filtered retained link list. Ready continuation arguments provide it. | |
| expected_links_hash | No | Optional retained-link-list version guard. Ready continuations supply it; mismatch returns links_changed without using the old index. |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
No arguments | |||