OCR extraction from scanned PDF documents (English)
OCR extraction from scanned PDF documents (English). Uses Tesseract to read text from image-based pages that have no embedded text layer. Returns structured text and metadata. Ideal for invoices, receipts, contracts, and legacy documents. $0.05 per extraction.
Answeringour last check, 2026-09-24
1 of 1checks answered this week
476 msmedian answer time
$0.05listed price per call
$0.05price it asked us
Paid test badge: not yet. The checks above are free: we call the tool without paying and read the payment request it sends back. The Verified badge needs paid calls whose answers match the promised output, and nobody can buy a badge.
Endpoint
GET https://visual.hugen.tokyo/visual/ocr
| Category | Image and media |
|---|---|
| Provider host | visual.hugen.tokyo |
| Networks | eip155:8453, solana:5eykt4UsFv8P8NJdTREpY1vzqKqZKvdp |
| Payment schemes | exact |
| Self-reported calls, 30 days | 3 from 3 payers (the provider's figure, not ours) |
Our checks, last 30 days
| Day | Result | HTTP | Asked | Time |
|---|---|---|---|---|
| 2026-09-24 | valid payment request | 402 | $0.05 | 476 ms |
Example input (from the provider)
{
"method": "GET",
"queryParams": {
"max_pages": "50",
"url": "https://example.com/scanned-invoice.pdf"
},
"type": "http"
}
Promised output schema (from the provider)
{
"properties": {
"input": {
"required": [
"method"
]
}
}
}