GPT-4o Transcribe Diarize
Transcription model that identifies who is speaking when.
Specs
- Context
- 16K
- Max output
- 2K
- Input price
- $2.5/ 1M tokens
- Output price
- $10/ 1M tokens
- Released
- —
Context
- Context window
- 16K
- Max input
- —
- Max output
- 2K
Pricing · USD / 1M tokens
- Input
- $2.5
- Output
- $10
- Cache read
- —
- Cache write
- —
Modalities
- Input
- AudioText
- Output
- Text
API
- API types
Reasoning
- Not published
Info
- Status
- Deprecated
- Released
- —
- Knowledge cutoff
- 2024-06-01
Features
- Not published
How to call
1 provider
openaimodel = gpt-4o-transcribe-diarize
audio
POSThttps://api.openai.com/v1/audio/transcriptions
JSON
Standard format
{
"id": "gpt-4o-transcribe-diarize",
"object": "model",
"created": null,
"owned_by": "openai",
"name": "GPT-4o Transcribe Diarize",
"api": {
"types": [
"audio"
]
},
"limits": {
"context": 16000,
"input": null,
"output": 2000
},
"modalities": {
"input": [
"audio",
"text"
],
"output": [
"text"
]
},
"reasoning": {
"supported": null,
"efforts": []
},
"pricing": {
"currency": "USD",
"unit": "1M_tokens",
"input": 2.5,
"output": 10,
"cache_read": null,
"cache_write": null
},
"features": [],
"info": {
"status": "deprecated",
"release_date": null,
"knowledge_cutoff": "2024-06-01",
"description": "Transcription model that identifies who is speaking when.",
"docs": "https://developers.openai.com/api/docs/models/gpt-4o-transcribe-diarize",
"verified_at": "2026-10-11"
}
}Official sources
4