Skip to main content
For developers requiring audio support, Infercom provides OpenAI’s Whisper large-v3 model, which enables real-time transcriptions and translations.

Whisper-Large-v3

  • Model: Whisper-Large-v3
  • Description: State-of-the-art automatic speech recognition (ASR) and translation model. Developed by OpenAI and trained on 5M+ hours of labeled audio. Excels in multilingual and zero-shot speech tasks across diverse domains.
  • Model ID: Whisper-Large-v3
  • Supported languages: Multilingual

Core capabilities

  • Transcribes and translates extended audio inputs (up to 25 MB).
  • Demonstrates high accuracy in speech recognition and translation tasks.
  • Provides OpenAI-compatible endpoints for transcriptions and translations.

Request parameters

Example usage