Recorded Overview
Stable
Batch transcription jobs from audio URLs.
Recorded Overview
Recorded is batch transcription: submit a public audio_url, process asynchronously, receive complete results.
For synchronous client-VAD clips see Segment Transcription. For live streams see Realtime.
How it works
Submit audio URL (POST /v1/audio/transcriptions/jobs)
→ Probe audio format and duration
→ Slice into time-based chunks
→ Concurrent transcription
→ Aggregate (merge into global timestamps)
→ Optional LLM review
→ Webhook callback (optional)
When to use
| Scenario | Suitable |
|---|---|
| Post-meeting processing | ✅ |
| Batch audio file transcription | ✅ |
| Need precise timestamps | ✅ |
| Need structured results | ✅ |
| Need subtitle files | ✅ |
| Short utterance after client VAD | ❌ use Segment |
| Live microphone stream | ❌ use Realtime |
vs. Realtime and Segment
| Dimension | Batch jobs | Segment HTTP | Realtime |
|---|---|---|---|
| Transport | HTTP async | HTTP sync | WebSocket |
| Input | Audio URL | Multipart file | PCM16LE stream |
| Latency | Minutes | ~≤1.6s | Milliseconds |
| Output | Complete JSON | Transcript JSON | Event stream |
| Segmentation | Server (time slice) | Client (VAD) | Server (VAD) |
| Success | 202 + poll | 200 | WS events |
Current capabilities
| Capability | Status |
|---|---|
| Full-file transcription | ✅ |
| Segment timestamps | ✅ |
| Confidence scores | ✅ |
| Translation | Coming soon |
| Webhook callback | ✅ |
| LLM review correction | ✅ (optional) |
| Speaker diarization | Coming soon |
| SRT/VTT subtitles | ✅ (generated from timestamps) |
Next steps
- Transcribe Audio — submit a job
- Batch Jobs API
- Segment API — client-VAD clips
- Timestamps & Speakers
Handling Interruptions & Silence
Silence, pauses, VAD end-of-speech behavior, and connection keepalive.
Transcribe Audio
Submit an audio URL for batch transcription.
Resources
