POST /map
Discover URLs for a website with sitemap-first, best-effort discovery.
POST
Discover URLs for a website. By default the endpoint checks
robots.txt for one or more sitemap
URLs, follows sitemap indexes, and collects page URLs from sitemap XML (including .xml.gz
files). Search-engine site: discovery is used only when no sitemap URLs are found.
Request parameters (JSON body)
Required
url(string): Base website URL to discover URLs for.
Optional
sitemap(string, default:include):include,skip,only.includeuses SERP discovery only when no sitemap URLs are found;onlydisables SERP entirely;skipskips robots/sitemap discovery and goes straight to SERP.includeSubdomains(boolean, default:true): Whether to include subdomains of the registrable domain.ignoreQueryParameters(boolean, default:true): Whether URLs with query parameters should be excluded.ignoreCache(boolean, default:false): Bypass cached discovery results and refresh.limit(integer, default:1000, max:10000): Maximum number of links to return.timeout(integer): Optional overall request timeout in milliseconds.
Responses
- 200:
successpluslinks— a best-effort flat list of discovered URLs, each with atypecategory (product-details,article,homepage,unknown, etc.) and an optionaltitle/descriptionwhen discovered via search. - 422: Invalid map request (e.g. malformed
urlor missing hostname).
Authorizations
Bearer authentication header of the form Bearer <token>, where <token> is your auth token.
Body
application/json
Base website URL to discover URLs for.
Example:
"https://www.amazon.com"
Sitemap handling. include uses SERP only when no sitemap URLs are found.
Available options:
skip, include, only Whether to include subdomains of the registrable domain.
Whether URLs with query parameters should be excluded.
Bypass Redis-backed map discovery caches and refresh discovery.
Maximum number of links to return.
Required range:
x <= 10000Examples:
1000
2500
Optional overall map request timeout in milliseconds.
Example:
10000
