Extract all text from a PDF
Extract all text from a PDF. Send as pdf_base64 (base64-encoded PDF, max ~10 MB decoded). Returns text (full concatenated text), pages array (per-page text + char_count), page_count, and metadata (title, author, creator). Encode with: Buffer.from(pdfBytes).toString('base64'). Ideal for RAG pipelines, document QA, or LLM ingestion.
Not tested: has real-world effectsour last check, 2026-10-10
0 of 0checks answered this week
n/amedian answer time
$0.0055listed price per call
n/aprice it asked us
Paid test badge: not yet. The checks above are free: we call the tool without paying and read the payment request it sends back. The Verified badge needs paid calls whose answers match the promised output, and nobody can buy a badge.
Endpoint
POST https://agentsvc.io/api/v1/proxy/pdf-extract
| Category | Search and research |
|---|---|
| Provider host | agentsvc.io |
| Networks | eip155:137, eip155:42161, eip155:8453, solana:5eykt4UsFv8P8NJdTREpY1vzqKqZKvdp |
| Payment schemes | exact |
| Self-reported calls, 30 days | 7 from 1 payers (the provider's figure, not ours) |
Our checks, last 30 days
We never call tools that send, buy, move money or file anything, not even without paying.
Example input (from the provider)
{
"body": {
"max_pages": 10,
"pdf_base64": "<base64 of a PDF>"
},
"bodyType": "json",
"method": "POST",
"type": "http"
}
Promised output schema (from the provider)
{
"$schema": "https://json-schema.org/draft/2020-12/schema",
"properties": {
"input": {
"additionalProperties": false,
"properties": {
"body": {
"properties": {
"max_pages": {
"default": 50,
"description": "Maximum number of pages to extract. Default: 50. Use to limit processing time for large PDFs.",
"type": "integer"
},
"pdf_base64": {
"description": "Base64-encoded PDF file content. Decode a PDF file to base64 and pass it here. Max ~10 MB (unencoded).",
"type": "string"
}
},
"required": [
"pdf_base64"
]
},
"bodyType": {
"enum": [
"json",
"form-data",
"text"
],
"type": "string"
},
"method": {
"enum": [
"POST",
"PUT",
"PATCH"
],
"type": "string"
},
"type": {
"const": "http",
"type": "string"
}
},
"required": [
"type",
"method",
"bodyType",
"body"
],
"type": "object"
},
"output": {
"properties": {
"example": {
"properties": {
"data": {
"properties": {
"extracted_at": {
"description": "ISO 8601 timestamp of extraction",
"format": "date-time",
"type": "string"
},
"file_size_bytes": {
"description": "Size of the decoded PDF in bytes",
"type": "integer"
},
"metadata": {
"description": "PDF document metadata (if available)",
"properties": {
"author": {
"type": "string"
},
"creation_date": {
"type": "string"
},
"creator": {
"type": "string"
},
"producer": {
"type": "string"
},
"subject": {
"type": "string"
},
"title": {
"type": "string"
}
},
"type": "object"
},
"page_count": {
"description": "Total number of pages in the PDF",
"type": "integer"
},
"pages": {
"description": "Per-page text content (first max_pages pages)",
"items": {
"properties": {
"char_count": {
"description": "Number of characters on this page",
"type": "integer"
},
"page": {
"description": "Page number (1-based)",
"type": "integer"
},
"text": {
"description": "Extracted text for this page",
"type": "string"
}
},
"type": "object"
},
"type": "array"
},
"text": {
"description": "Full extracted text from all pages, joined with newlines. Preserves paragraph structure where possible.",
"type": "string"
}
},
"required": [
"text",
"page_count",
"extracted_at"
],
"type": "object"
},
"success": {
"type": "boolean"
}
},
"required": [
"success",
"data"
],
"type": "object"
},
"type": {
"type": "string"
}
},
"required": [
"type"
],
"type": "object"
}
},
"required": [
"input"
],
"type": "object"
}