Crawl a website from a starting URL and return each page as clean markdown
Crawl a website from a starting URL and return each page as clean markdown: breadth-first over internal links, bounded by page count and depth, honouring robots.txt, with per-page title, status, depth and outbound links. Use it when an agent needs a whole section of a site rather than one known page.
Answeringour last check, 2026-09-24
1 of 1checks answered this week
948 msmedian answer time
$0.02listed price per call
$0.02price it asked us
Paid test badge: not yet. The checks above are free: we call the tool without paying and read the payment request it sends back. The Verified badge needs paid calls whose answers match the promised output, and nobody can buy a badge.
Endpoint
POST https://agent402.tools/api/site-crawl
| Category | Search and research |
|---|---|
| Provider host | agent402.tools |
| Networks | algorand:wGHE2Pwdvd7S12BL5FaOP20EGYesN73ktiC1qzkkit8=, eip155:10, eip155:1329, eip155:137, eip155:143, eip155:42161, eip155:42220, eip155:43114, eip155:4663, eip155:8453, solana:5eykt4UsFv8P8NJdTREpY1vzqKqZKvdp, stellar:pubnet |
| Payment schemes | exact, upto |
| Self-reported calls, 30 days | 3 from 1 payers (the provider's figure, not ours) |
Our checks, last 30 days
| Day | Result | HTTP | Asked | Time |
|---|---|---|---|---|
| 2026-09-24 | valid payment request | 402 | $0.02 | 948 ms |
Example input (from the provider)
{
"body": {
"limit": 3,
"maxDepth": 1,
"url": "https://example.com"
},
"bodyType": "json",
"method": "POST",
"type": "http"
}
Promised output schema (from the provider)
{
"$schema": "https://json-schema.org/draft/2020-12/schema",
"properties": {
"input": {
"additionalProperties": false,
"properties": {
"body": {
"properties": {
"excludePatterns": {
"description": "Never follow links whose URL contains any of these substrings (max 20)",
"items": {
"type": "string"
},
"type": "array"
},
"format": {
"description": "Page content format (default markdown)",
"enum": [
"markdown",
"text"
],
"type": "string"
},
"includePatterns": {
"description": "Only follow links whose URL contains at least one of these substrings (max 20)",
"items": {
"type": "string"
},
"type": "array"
},
"limit": {
"description": "Max pages to fetch, 1-20 (default 10); failed fetches count toward it",
"type": "integer"
},
"maxCharsPerPage": {
"description": "Cap on content characters per page, 200-20000 (default 8000)",
"type": "integer"
},
"maxDepth": {
"description": "Link depth from the start URL, 0-2 (default 1)",
"type": "integer"
},
"sameHost": {
"description": "true (default): stay on the start host (www and bare host count as one); false: also follow subdomains of the start site",
"type": "boolean"
},
"url": {
"description": "Start URL",
"type": "string"
}
},
"required": [
"url"
]
},
"bodyType": {
"enum": [
"json",
"form-data",
"text"
],
"type": "string"
},
"method": {
"enum": [
"POST"
],
"type": "string"
},
"type": {
"const": "http",
"type": "string"
}
},
"required": [
"type",
"method",
"bodyType",
"body"
],
"type": "object"
},
"output": {
"properties": {
"example": {
"type": "object"
},
"type": {
"type": "string"
}
},
"required": [
"type"
],
"type": "object"
}
},
"required": [
"input"
],
"type": "object"
}