Web scraper, URL to Markdown
Web scraper, URL to Markdown. Default tool for reading ONE web page whose URL you have: returns clean Markdown (title, meta description, main content with headings, lists and links). The fastest and cheapest way to give an LLM the content of a page. If a JavaScript app comes back nearly empty, set render:true (real browser, slower, plain text) or use webpage-reader. For 2 to 5 URLs use web-read-batch, for a whole site site-crawl, and if you have a question but no URL search-read.
Paid test badge: not yet. The checks above are free: we call the tool without paying and read the payment request it sends back. The Verified badge needs paid calls whose answers match the promised output, and nobody can buy a badge.
Endpoint
POST https://agentsvc.io/api/v1/proxy/web-scrape
| Category | Code and developer |
|---|---|
| Provider host | agentsvc.io |
| Networks | eip155:137, eip155:42161, eip155:8453, solana:5eykt4UsFv8P8NJdTREpY1vzqKqZKvdp |
| Payment schemes | exact |
| Self-reported calls, 30 days | 11 from 3 payers (the provider's figure, not ours) |
Our checks, last 30 days
| Day | Result | HTTP | Asked | Time |
|---|---|---|---|---|
| 2026-10-10 | valid payment request | 402 | $0.0025 | 891 ms |
Example input (from the provider)
{
"body": {
"max_chars": 8000,
"url": "https://example.com"
},
"bodyType": "json",
"method": "POST",
"type": "http"
}
Promised output schema (from the provider)
{
"$schema": "https://json-schema.org/draft/2020-12/schema",
"properties": {
"input": {
"additionalProperties": false,
"properties": {
"body": {
"properties": {
"include_links": {
"default": false,
"description": "Also return up to 100 absolute links found on the page",
"type": "boolean"
},
"max_chars": {
"default": 8000,
"description": "Truncate Markdown after this many characters",
"maximum": 60000,
"type": "integer"
},
"render": {
"default": false,
"description": "Render with a headless browser (for JavaScript-heavy pages). Returns plain text instead of Markdown.",
"type": "boolean"
},
"url": {
"description": "Public http(s) URL to scrape",
"format": "uri",
"type": "string"
}
},
"required": [
"url"
]
},
"bodyType": {
"enum": [
"json",
"form-data",
"text"
],
"type": "string"
},
"method": {
"enum": [
"POST",
"PUT",
"PATCH"
],
"type": "string"
},
"type": {
"const": "http",
"type": "string"
}
},
"required": [
"type",
"method",
"bodyType",
"body"
],
"type": "object"
},
"output": {
"properties": {
"example": {
"properties": {
"data": {
"properties": {
"description": {
"description": "Meta description if present",
"type": "string"
},
"fetch_ms": {
"type": "integer"
},
"fetched_at": {
"format": "date-time",
"type": "string"
},
"links": {
"items": {
"properties": {
"href": {
"type": "string"
},
"text": {
"type": "string"
}
},
"type": "object"
},
"type": "array"
},
"markdown": {
"description": "Page content as Markdown",
"type": "string"
},
"rendered": {
"description": "True if a headless browser was used",
"type": "boolean"
},
"status": {
"description": "HTTP status of the fetched page",
"type": "integer"
},
"title": {
"type": "string"
},
"truncated": {
"type": "boolean"
},
"url": {
"description": "Final URL after redirects",
"type": "string"
},
"word_count": {
"type": "integer"
}
},
"required": [
"url",
"status",
"title",
"markdown",
"word_count",
"truncated",
"fetched_at"
],
"type": "object"
},
"success": {
"type": "boolean"
}
},
"required": [
"success",
"data"
],
"type": "object"
},
"type": {
"type": "string"
}
},
"required": [
"type"
],
"type": "object"
}
},
"required": [
"input"
],
"type": "object"
}