Skip to main content
Use the id value as the model parameter in API requests. 5 models currently available.

Usage

Speech-to-text models transcribe spoken audio into written text. They are accessed via the Audio Transcriptions API.

Supported audio formats

mp3, mp4, mpeg, mpga, m4a, wav, webm, flac, ogg

Response formats

Pricing is billed per second of input audio. See the Audio Transcriptions API for request examples and parameter details.