Speech-to-text transcription for voice messages and audio clips
Speech-to-text transcription for voice messages and audio clips. Upload a wav, mp3, m4a, ogg or webm file as multipart form-data and get the spoken text back as JSON. Runs a dedicated ASR model, so short noisy voice notes transcribe accurately. Accepts clips up to 10 minutes / 25 MB, billed per 10-second block. Use it to turn voice notes, call recordings, podcast clips or voicemail into text an agent can read.
Answeringour last check, 2026-09-24
1 of 1checks answered this week
11129 msmedian answer time
$0.015listed price per call
$0.015price it asked us
Paid test badge: not yet. The checks above are free: we call the tool without paying and read the payment request it sends back. The Verified badge needs paid calls whose answers match the promised output, and nobody can buy a badge.
Endpoint
POST https://voice.cyberwarex.com/transcribe
| Category | Image and media |
|---|---|
| Provider host | voice.cyberwarex.com |
| Networks | eip155:8453 |
| Payment schemes | exact |
| Self-reported calls, 30 days | 7 from 1 payers (the provider's figure, not ours) |
Our checks, last 30 days
| Day | Result | HTTP | Asked | Time |
|---|---|---|---|---|
| 2026-09-24 | valid payment request | 402 | $0.015 | 11129 ms |
Example input (from the provider)
{
"body": {
"file": "<multipart audio file: voice-note.mp3>"
},
"bodyType": "form-data",
"method": "POST",
"type": "http"
}
Promised output schema (from the provider)
{
"$schema": "https://json-schema.org/draft/2020-12/schema",
"properties": {
"input": {
"additionalProperties": false,
"properties": {
"body": {
"properties": {
"file": {
"description": "Audio file to transcribe (wav, mp3, m4a, ogg, webm, flac). Max 10 minutes and 25 MB.",
"format": "binary",
"type": "string"
}
},
"required": [
"file"
],
"type": "object"
},
"bodyType": {
"enum": [
"json",
"form-data",
"text"
],
"type": "string"
},
"method": {
"enum": [
"POST",
"PUT",
"PATCH"
],
"type": "string"
},
"type": {
"const": "http",
"type": "string"
}
},
"required": [
"type",
"method",
"bodyType",
"body"
],
"type": "object"
},
"output": {
"properties": {
"example": {
"properties": {
"charged": {
"description": "False when silence was detected and no payment was taken.",
"type": "boolean"
},
"note": {
"description": "Set to 'no speech detected' when the clip is silent.",
"type": "string"
},
"text": {
"description": "Transcribed speech from the uploaded audio.",
"type": "string"
}
},
"type": "object"
},
"type": {
"type": "string"
}
},
"required": [
"type"
],
"type": "object"
}
},
"required": [
"input"
],
"type": "object"
}