
Batch speech recognition model with accurate transcription in 90+ languages, keyterm prompting, entity detection, word-level timestamps, speaker diarization, audio tagging, and language detection.
MODALITIES
CONTEXT
Not specified
INPUT PRICE
$0.0037
per minute
PROVIDERS / REGIONS
All live provider offers for this model, with region and token pricing.
Provider
Product
Context
Location
Input/min
Cached input
elevenlabs
Scribe V2
Not specified
us
$0.0037
0
Additional provider routing data is not available yet.
CAPABILITIES
Supported inputs, outputs, and model features from the live CMS record.
Audio input
Audio modality
Speech to text input
RELATED MODELS
Other live models from the same provider.

