Speech transcription model for accurate audio-to-text and captioning workflows
Details
| Field | Value |
|---|---|
| id | openai/whisper-large-v3 |
| family | whisper |
| modality | audio->text |
| context_length | 448 |
| max_output_tokens | 448 |
| release_date | 2023-11-06 |
| is_open_weights | yes |
| supports_tool_call | no |
| supports_reasoning | no |
| supports_structured_output | no |
| supports_attachment | no |
| input_modalities | audio |
| output_modalities | text |