ToolAssay

Extract all text from a PDF

Extract all text from a PDF. Send as pdf_base64 (base64-encoded PDF, max ~10 MB decoded). Returns text (full concatenated text), pages array (per-page text + char_count), page_count, and metadata (title, author, creator). Encode with: Buffer.from(pdfBytes).toString('base64'). Ideal for RAG pipelines, document QA, or LLM ingestion.

Not tested: has real-world effectsour last check, 2026-10-10
0 of 0checks answered this week
n/amedian answer time
$0.0055listed price per call
n/aprice it asked us

Paid test badge: not yet. The checks above are free: we call the tool without paying and read the payment request it sends back. The Verified badge needs paid calls whose answers match the promised output, and nobody can buy a badge.

Endpoint

POST https://agentsvc.io/api/v1/proxy/pdf-extract

CategorySearch and research
Provider hostagentsvc.io
Networkseip155:137, eip155:42161, eip155:8453, solana:5eykt4UsFv8P8NJdTREpY1vzqKqZKvdp
Payment schemesexact
Self-reported calls, 30 days7 from 1 payers (the provider's figure, not ours)

Our checks, last 30 days

We never call tools that send, buy, move money or file anything, not even without paying.

Example input (from the provider)

{
  "body": {
    "max_pages": 10,
    "pdf_base64": "<base64 of a PDF>"
  },
  "bodyType": "json",
  "method": "POST",
  "type": "http"
}

Promised output schema (from the provider)

{
  "$schema": "https://json-schema.org/draft/2020-12/schema",
  "properties": {
    "input": {
      "additionalProperties": false,
      "properties": {
        "body": {
          "properties": {
            "max_pages": {
              "default": 50,
              "description": "Maximum number of pages to extract. Default: 50. Use to limit processing time for large PDFs.",
              "type": "integer"
            },
            "pdf_base64": {
              "description": "Base64-encoded PDF file content. Decode a PDF file to base64 and pass it here. Max ~10 MB (unencoded).",
              "type": "string"
            }
          },
          "required": [
            "pdf_base64"
          ]
        },
        "bodyType": {
          "enum": [
            "json",
            "form-data",
            "text"
          ],
          "type": "string"
        },
        "method": {
          "enum": [
            "POST",
            "PUT",
            "PATCH"
          ],
          "type": "string"
        },
        "type": {
          "const": "http",
          "type": "string"
        }
      },
      "required": [
        "type",
        "method",
        "bodyType",
        "body"
      ],
      "type": "object"
    },
    "output": {
      "properties": {
        "example": {
          "properties": {
            "data": {
              "properties": {
                "extracted_at": {
                  "description": "ISO 8601 timestamp of extraction",
                  "format": "date-time",
                  "type": "string"
                },
                "file_size_bytes": {
                  "description": "Size of the decoded PDF in bytes",
                  "type": "integer"
                },
                "metadata": {
                  "description": "PDF document metadata (if available)",
                  "properties": {
                    "author": {
                      "type": "string"
                    },
                    "creation_date": {
                      "type": "string"
                    },
                    "creator": {
                      "type": "string"
                    },
                    "producer": {
                      "type": "string"
                    },
                    "subject": {
                      "type": "string"
                    },
                    "title": {
                      "type": "string"
                    }
                  },
                  "type": "object"
                },
                "page_count": {
                  "description": "Total number of pages in the PDF",
                  "type": "integer"
                },
                "pages": {
                  "description": "Per-page text content (first max_pages pages)",
                  "items": {
                    "properties": {
                      "char_count": {
                        "description": "Number of characters on this page",
                        "type": "integer"
                      },
                      "page": {
                        "description": "Page number (1-based)",
                        "type": "integer"
                      },
                      "text": {
                        "description": "Extracted text for this page",
                        "type": "string"
                      }
                    },
                    "type": "object"
                  },
                  "type": "array"
                },
                "text": {
                  "description": "Full extracted text from all pages, joined with newlines. Preserves paragraph structure where possible.",
                  "type": "string"
                }
              },
              "required": [
                "text",
                "page_count",
                "extracted_at"
              ],
              "type": "object"
            },
            "success": {
              "type": "boolean"
            }
          },
          "required": [
            "success",
            "data"
          ],
          "type": "object"
        },
        "type": {
          "type": "string"
        }
      },
      "required": [
        "type"
      ],
      "type": "object"
    }
  },
  "required": [
    "input"
  ],
  "type": "object"
}

This page as JSON