Use the
id value as the model parameter in API requests. 5 models currently available.Usage
Speech-to-text models transcribe spoken audio into written text. They are accessed via the Audio Transcriptions API.Supported audio formats
mp3, mp4, mpeg, mpga, m4a, wav, webm, flac, ogg
Response formats
Pricing is billed per second of input audio. See the Audio Transcriptions API for request examples and parameter details.