Recorded Overview

Stable
Batch transcription jobs from audio URLs.

Recorded Overview

Recorded is batch transcription: submit a public audio_url, process asynchronously, receive complete results.

For synchronous client-VAD clips see Segment Transcription. For live streams see Realtime.

How it works

Submit audio URL (POST /v1/audio/transcriptions/jobs)
  → Probe audio format and duration
  → Slice into time-based chunks
  → Concurrent transcription
  → Aggregate (merge into global timestamps)
  → Optional LLM review
  → Webhook callback (optional)

When to use

ScenarioSuitable
Post-meeting processing✅
Batch audio file transcription✅
Need precise timestamps✅
Need structured results✅
Need subtitle files✅
Short utterance after client VAD❌ use Segment
Live microphone stream❌ use Realtime

vs. Realtime and Segment

DimensionBatch jobsSegment HTTPRealtime
TransportHTTP asyncHTTP syncWebSocket
InputAudio URLMultipart filePCM16LE stream
LatencyMinutes~≤1.6sMilliseconds
OutputComplete JSONTranscript JSONEvent stream
SegmentationServer (time slice)Client (VAD)Server (VAD)
Success202 + poll200WS events

Current capabilities

CapabilityStatus
Full-file transcription✅
Segment timestamps✅
Confidence scores✅
TranslationComing soon
Webhook callback✅
LLM review correction✅ (optional)
Speaker diarizationComing soon
SRT/VTT subtitles✅ (generated from timestamps)

Next steps