Text to Speech (TTS) Models
Lightning v2
Legacy An upgrade from the Lightning Large model, offering improved performance and
quality. It supports 16 languages, making it suitable for a wider range of
applications requiring expressive and high-quality speech synthesis.
Lightning v3.1
Latest Release A 44 kHz model delivering natural, expressive, and realistic speech. Supports voice cloning with ultra-low latency. 15 languages with auto-detection and mid-sentence switching.
Speech to Text (STT) Models
Pulse STT
Low-latency speech recognition for real-time and pre-recorded transcription.
Automatic language detection across 39 languages.
Geo-location Based Routing
Waves intelligently routes every request to the nearest server cluster to ensure the lowest possible latency for your applications. We currently operate server clusters in:- 🇮🇳 India (Mumbai)
- 🇺🇸 USA (Oregon)
Model Overview (TTS)
Model Overview (STT)
Note: The API uses ISO 639-1 language codes - Set
1 (2-letter
codes) to specify supported languages.

