ToolAssay

Web scraping API

Web scraping API: fetch any public web page and get its readable content as clean markdown — title, author, canonical URL, boilerplate stripped. Honest User-Agent, robots.txt honored (explicit Disallow returns an unpaid 403), private/internal targets refused, at most 3 safety-revalidated redirects, 1MB input / 100k character output caps. HTML pages only. Use for research agents, content extraction, summarization pipelines, and RAG ingestion. Cached up to 5 minutes per URL.

Answeringour last check, 2026-09-24
1 of 1checks answered this week
21 msmedian answer time
$0.03listed price per call
$0.03price it asked us

Paid test badge: not yet. The checks above are free: we call the tool without paying and read the payment request it sends back. The Verified badge needs paid calls whose answers match the promised output, and nobody can buy a badge.

Endpoint

GET https://gateway.stride20k.com/web/extract

CategorySearch and research
Provider hostgateway.stride20k.com
Networkseip155:8453
Payment schemesexact
Self-reported calls, 30 days1 from 1 payers (the provider's figure, not ours)

Our checks, last 30 days

DayResultHTTPAskedTime
2026-09-24 valid payment request 402$0.03 21 ms

Example input (from the provider)

{
  "method": "GET",
  "queryParams": {
    "url": "https://example.com/"
  },
  "type": "http"
}

Promised output schema (from the provider)

{
  "$schema": "https://json-schema.org/draft/2020-12/schema",
  "properties": {
    "input": {
      "additionalProperties": false,
      "properties": {
        "method": {
          "enum": [
            "GET"
          ],
          "type": "string"
        },
        "queryParams": {
          "properties": {
            "url": {
              "description": "Absolute http(s) URL of the page to extract.",
              "pattern": "https?://.+",
              "type": "string",
              "urlSafety": true
            }
          },
          "required": [
            "url"
          ],
          "type": "object"
        },
        "type": {
          "const": "http",
          "type": "string"
        }
      },
      "required": [
        "type",
        "method"
      ],
      "type": "object"
    },
    "output": {
      "properties": {
        "example": {
          "properties": {
            "data": {
              "properties": {
                "author": {
                  "description": "meta author, or null.",
                  "type": [
                    "string",
                    "null"
                  ]
                },
                "canonicalUrl": {
                  "description": "rel=canonical link if the page declares one, else null.",
                  "type": [
                    "string",
                    "null"
                  ]
                },
                "charCount": {
                  "description": "Length of the markdown field.",
                  "type": "integer"
                },
                "finalUrl": {
                  "description": "URL that actually served the content.",
                  "type": "string"
                },
                "markdown": {
                  "description": "Extracted readable content as markdown.",
                  "minLength": 1,
                  "type": "string"
                },
                "title": {
                  "description": "Document title, or null.",
                  "type": [
                    "string",
                    "null"
                  ]
                },
                "truncated": {
                  "description": "True if output hit the 100k-char cap.",
                  "type": "boolean"
                },
                "url": {
                  "description": "The URL requested (after redirect re-validation).",
                  "type": "string"
                }
              },
              "required": [
                "url",
                "markdown",
                "truncated"
              ],
              "type": "object"
            },
            "endpoint": {
              "description": "Always \"/web/extract\".",
              "type": "string"
            },
            "meta": {
              "properties": {
                "attribution": {
                  "description": "Licensing attribution when the source requires it.",
                  "type": [
                    "string",
                    "null"
                  ]
                },
                "cached": {
                  "description": "Whether this response was served from cache.",
                  "type": "boolean"
                },
                "fetchedAt": {
                  "description": "ISO 8601 time the data was actually retrieved from the upstream.",
                  "type": "string"
                },
                "source": {
                  "description": "Upstream source: direct fetch (buyer-directed).",
                  "type": "string"
                }
              },
              "required": [
                "source",
                "cached",
                "fetchedAt"
              ],
              "type": "object"
            },
            "ok": {
              "description": "true on success.",
              "type": "boolean"
            }
          },
          "required": [
            "ok",
            "endpoint",
            "data",
            "meta"
          ],
          "type": "object"
        },
        "type": {
          "type": "string"
        }
      },
      "required": [
        "type"
      ],
      "type": "object"
    }
  },
  "required": [
    "input"
  ],
  "type": "object"
}

This page as JSON