AICredits logo
Capability directory8 published models

Audio Models

Browse speech and transcription models available through AICredits.

Audio models cover speech generation and transcription workflows. Check each model detail page for supported endpoints.

google/lyria-3-clip-preview

30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate...

Transcriptions APIVisionTranscription
Context
1.0M
Input
Per minute pricing
View details
google/lyria-3-pro-preview

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz...

Transcriptions APIVisionTranscription
Context
1.0M
Input
Per minute pricing
View details
mistralai/voxtral-small-24b-2507

Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance. It excels at speech transcription, translation and audio understanding. Input audio...

Transcriptions APITranscription
Context
32K
Input
Per minute pricing
Cached input
Cached ₹1.00/1M
View details
openai/whisper-1

OpenAI speech-to-text transcription model.

Transcriptions APITranscription
Context
Unknown
Input
Per minute pricing
View details