ToolAssay

Extract text from a web page or PDF as clean Markdown — HTML to Markdown for any

Extract text from a web page or PDF as clean Markdown — HTML to Markdown for any URL: strips scripts, nav, ads, and boilerplate while preserving headings, links, lists, tables, code blocks, and blockquotes; extracts the text layer from PDFs. Returns Markdown body, title, word count, and a quality grade. Web scraper / reader for article and main content. For JS-rendered or bot-walled pages a plain fetch can't read, use /exa/contents.

Answeringour last check, 2026-09-24
1 of 1checks answered this week
1140 msmedian answer time
$0.003listed price per call
$0.003price it asked us

Paid test badge: not yet. The checks above are free: we call the tool without paying and read the payment request it sends back. The Verified badge needs paid calls whose answers match the promised output, and nobody can buy a badge.

Endpoint

GET https://netintel.dev/web/extract

CategoryCode and developer
Provider hostnetintel.dev
Networkseip155:8453, solana:5eykt4UsFv8P8NJdTREpY1vzqKqZKvdp
Payment schemesexact
Self-reported calls, 30 days350 from 8 payers (the provider's figure, not ours)

Our checks, last 30 days

DayResultHTTPAskedTime
2026-09-24 valid payment request 402$0.003 1140 ms

Example input (from the provider)

{
  "method": "GET",
  "queryParams": {
    "url": "https://www.sitemaps.org/protocol.html"
  },
  "type": "http"
}

Promised output schema (from the provider)

{
  "$schema": "https://json-schema.org/draft/2020-12/schema",
  "properties": {
    "input": {
      "additionalProperties": false,
      "properties": {
        "method": {
          "enum": [
            "GET"
          ],
          "type": "string"
        },
        "queryParams": {
          "properties": {
            "url": {
              "description": "Public URL of an HTML page or PDF to extract (e.g. https://www.sitemaps.org/protocol.html)",
              "type": "string"
            }
          },
          "required": [
            "url"
          ],
          "type": "object"
        },
        "type": {
          "const": "http",
          "type": "string"
        }
      },
      "required": [
        "type",
        "method"
      ],
      "type": "object"
    },
    "output": {
      "properties": {
        "example": {
          "properties": {
            "char_count": {
              "type": "number"
            },
            "content_type": {
              "description": "article | pdf | other (non-HTML/PDF falls back to best-effort plain text)",
              "type": "string"
            },
            "final_url": {
              "type": "string"
            },
            "findings": {
              "type": "array"
            },
            "grade": {
              "type": "string"
            },
            "markdown": {
              "type": "string"
            },
            "output_bytes": {
              "type": "number"
            },
            "score": {
              "type": "number"
            },
            "status_code": {
              "type": "number"
            },
            "title": {
              "type": "string"
            },
            "truncated": {
              "type": "boolean"
            },
            "url": {
              "type": "string"
            },
            "word_count": {
              "type": "number"
            }
          },
          "type": "object"
        },
        "type": {
          "type": "string"
        }
      },
      "required": [
        "type"
      ],
      "type": "object"
    }
  },
  "required": [
    "input"
  ],
  "type": "object"
}

This page as JSON