ModelInfo
English
All models

Scribe v2

Batch speech-to-text with word-level timestamps, speaker diarization and keyterm prompting.

ElevenLabsscribe_v2Report incorrect data

Specs

Context
—
Max output
—
Input price
$0.22/ hour
Output price
—
Released
—

Context

Not published

Pricing · USD / hour

Input
$0.22
Output
—
Cache read
—
Cache write
—

Modalities

Input
Audio
Output
Text

API

API types
audio

Reasoning

Not published

Info

Status
Active
Released
—
Knowledge cutoff
—

Features

Not published
Source · official docsVerified 2026-10-11

How to call

1 provider

elevenlabsmodel = scribe_v2

audio
POSThttps://api.elevenlabs.io/v1/speech-to-text

JSON

Standard format
scribe_v2.json
{
  "id": "scribe_v2",
  "object": "model",
  "created": null,
  "owned_by": "elevenlabs",
  "name": "Scribe v2",
  "api": {
    "types": [
      "audio"
    ]
  },
  "limits": {
    "context": null,
    "input": null,
    "output": null
  },
  "modalities": {
    "input": [
      "audio"
    ],
    "output": [
      "text"
    ]
  },
  "reasoning": {
    "supported": null,
    "efforts": []
  },
  "pricing": {
    "currency": "USD",
    "unit": "hour",
    "input": 0.22,
    "output": null,
    "cache_read": null,
    "cache_write": null
  },
  "features": [],
  "info": {
    "status": "active",
    "release_date": null,
    "knowledge_cutoff": null,
    "description": "Batch speech-to-text with word-level timestamps, speaker diarization and keyterm prompting.",
    "docs": "https://elevenlabs.io/docs/overview/models",
    "verified_at": "2026-10-11"
  }
}

Official sources

3