For the complete documentation index, see llms.txt. This page is also available as Markdown.

Audio Models

Here are all the audio models currently available in FLORA.

Speech & Voice

These models generate or transcribe speech.

Model
Est. time (s)
Description
Best for

ElevenLabs Multilingual v2

15

Natural text-to-speech with multilingual voice synthesis.

Voiceovers, narration, multilingual content, accessible media.

ElevenLabs Scribe v2

10

Transcribe speech to text.

Transcription, subtitles, meeting notes, content indexing.

Gemini 3.1 Flash TTS

10

Text to speech with prompted expressiveness.

Expressive voiceovers, character dialog, and stylized narration.

Sound Effects & Music

These models generate sound effects and music from text prompts.

Model
Est. time (s)
Description
Best for

ElevenLabs Music v1

30

Generate music using a text prompt.

Background tracks, musical stings, and prompt-driven composition.

ElevenLabs Sound Effects

15

Generate sound effects and Foley from text descriptions.

Foley, ambient soundscapes, UI sounds, game audio.

Last updated

Was this helpful?