elevenlabs

scribe_v2

Batch speech recognition model with accurate transcription in 90+ languages, keyterm prompting, entity detection, word-level timestamps, speaker diarization, audio tagging, and language detection.

Provider:

elevenlabs

Model type:

stt

Location:

us

Context Window

Intelligence Rating

Speed Rating

Cost Efficiency Rating

Pricing

$

0

Input tokens per million

$

0

Output tokens per million

Features

Audio Input

Supported

Create an account and start building today.

Create an account and start building today.