EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

text to speech

Page 1 · 28 per page
text-to-speech
MiniMaxREVIEW REQUIRED

MiniMax Speech-02 HD

fal-ai/minimax/speech-02-hd

Generate speech from text prompts and different voices using the MiniMax Speech-02 HD model, which leverages advanced AI techniques to create high-quality text-to-speech.

speech
text-to-speech
GoogleREVIEW REQUIRED

Gemini 3.1 Flash Tts

fal-ai/gemini-3.1-flash-tts

Newest audio model from Google introduces granular audio tags that give you precise control to direct AI speech for expressive audio generation.

lipsyncavatar
text-to-speech
ElevenLabsREVIEW REQUIRED

ElevenLabs TTS Turbo v2.5

fal-ai/elevenlabs/tts/turbo-v2.5

Generate high-speed text-to-speech audio using ElevenLabs TTS Turbo v2.5.

audio
text-to-speech
MiniMaxREVIEW REQUIRED

MiniMax Speech 2.8 [HD]

fal-ai/minimax/speech-2.8-hd

Generate speech from text prompts and different voices using the MiniMax Speech-2.8 HD model, which leverages advanced AI techniques to create high-quality text-to-speech.

text-to-speech
MiniMaxREVIEW REQUIRED

MiniMax Voice Cloning

fal-ai/minimax/voice-clone

Clone a voice from a sample audio and generate speech from text prompts using the MiniMax model, which leverages advanced AI techniques to create high-quality text-to-speech.

speech
text-to-speech
MiniMaxREVIEW REQUIRED

MiniMax Speech-02 Turbo

fal-ai/minimax/speech-02-turbo

Generate fast speech from text prompts and different voices using the MiniMax Speech-02 Turbo model, which leverages advanced AI techniques to create high-quality text-to-speech.

speech
text-to-speech
AlibabaREVIEW REQUIRED

Qwen 3 TTS - Text to Speech [1.7B]

fal-ai/qwen-3-tts/text-to-speech/1.7b

Bring speech to your texts using Qwen3-TTS Custom-Voice model with pre-trained voices or use your custom voice with Qwen3-TTS Clone Voice model

text-to-speech
text-to-speech
MiniMaxREVIEW REQUIRED

MiniMax Speech 2.8 [Turbo]

fal-ai/minimax/speech-2.8-turbo

Generate speech from text prompts and different voices using the MiniMax Speech-2.8 Turbo model, which leverages advanced AI techniques to create high-quality text-to-speech.

text-to-speech
xAIREVIEW REQUIRED

xAI Text to Speech

xai/tts/v1

Generate speech with expressive and realistic voices from xAI

text-to-speech
falREVIEW REQUIRED

Inworld TTS-1.5 Max

fal-ai/inworld-tts

Text to Speech Endpoint for Inworld's TTS-1.5 Max.

text-to-speechinworldtts
text-to-speech
MiniMaxREVIEW REQUIRED

MiniMax Voice Design

fal-ai/minimax/voice-design

Design a personalized voice from a text description, and generate speech from text prompts using the MiniMax model, which leverages advanced AI techniques to create high-quality text-to-speech.

speech
text-to-speech
falREVIEW REQUIRED

Chatterbox

fal-ai/chatterbox/text-to-speech

Whether you're working on memes, videos, games, or AI agents, Chatterbox brings your content to life. Use the first tts from resemble ai.

text-to-speech
text-to-speech
AlibabaREVIEW REQUIRED

Qwen 3 TTS - Voice Design [1.7B]

fal-ai/qwen-3-tts/voice-design/1.7b

Create custom voices using Qwen3-TTS Voice Design model and later use Clone Voice model to create your own voices!

text-to-speechvoice-design
text-to-speech
falREVIEW REQUIRED

Index TTS 2.0

fal-ai/index-tts-2/text-to-speech

Generate natural, clear speeches using Index TTS 2.0 from IndexTeam

text-to-speech
text-to-speech
MiniMaxREVIEW REQUIRED

MiniMax Speech 2.6 [HD]

fal-ai/minimax/speech-2.6-hd

Generate speech from text prompts and different voices using the MiniMax Speech-2.6 HD model, which leverages advanced AI techniques to create high-quality text-to-speech.

text-to-speech
text-to-speech
ByteDanceREVIEW REQUIRED

Bytedance Seed Speech Text to Speech

fal-ai/bytedance/seed-speech/tts/v2

Seed Speech developed by ByteDance, is a family of large-scale text-to-speech models capable of synthesizing speech that is virtually indistinguishable from human speech.

stylizedtransformlipsync
text-to-speech
resemble-aiREVIEW REQUIRED

Chatterboxhd

resemble-ai/chatterboxhd/text-to-speech

Generate expressive, natural speech with Resemble AI's Chatterbox. Features unique emotion control, instant voice cloning from short audio, and built-in watermarking.

text-to-speech
MiniMaxREVIEW REQUIRED

MiniMax Speech 2.6 [Turbo]

fal-ai/minimax/speech-2.6-turbo

Generate speech from text prompts and different voices using the MiniMax Speech-2.6 HD model, which leverages advanced AI techniques to create high-quality text-to-speech.

text-to-speech
text-to-speech
falREVIEW REQUIRED

Chatterbox

fal-ai/chatterbox/text-to-speech/multilingual

Whether you're working on memes, videos, games, or AI agents, Chatterbox brings your content to life. Use the first tts from resemble ai.

text-to-speechmultilingual
text-to-speech
AlibabaREVIEW REQUIRED

Qwen Audio 3.0 TTS (Flash)

alibaba/qwen-audio-3-tts

Generate natural multilingual speech from text with fast voice and language control using Qwen Audio 3.0 TTS Flash.

text-to-speechaudiospeech-synthesismultilingual
text-to-speech
falREVIEW REQUIRED

Zonos2 Text to Speech

fal-ai/zonos2

Zonos2 is a text-to-speech model that clones a voice from a short sample and speaks naturally across many languages.

text-to-speechttsvoice cloning
text-to-speech
falREVIEW REQUIRED

VibeVoice 7B

fal-ai/vibevoice/7b

Generate long, expressive multi-voice speech using Microsoft's powerful TTS

text-to-speechmulti-speakerpodcast
text-to-speech
falREVIEW REQUIRED

Maya1

fal-ai/maya

Maya1 is a state-of-the-art speech model by Maya Research for expressive voice generation, built to capture real human emotion and precise voice design.

text-to-speechtts
text-to-speech
falREVIEW REQUIRED

Orpheus TTS

fal-ai/orpheus-tts

Orpheus TTS is a state-of-the-art, Llama-based Speech-LLM designed for high-quality, empathetic text-to-speech generation. This model has been finetuned to deliver human-level speech synthesis, achieving exceptional clarity, expressiveness, and real-time performances.

text to speechvoice synthesishigh-fidelity
text-to-speech
falREVIEW REQUIRED

Dia

fal-ai/dia-tts

Dia directly generates realistic dialogue from transcripts. Audio conditioning enables emotion control. Produces natural nonverbals like laughter and throat clearing.

text-to-speech
text-to-speech
KlingREVIEW REQUIRED

Kling TTS

fal-ai/kling-video/v1/tts

Generate speech from text prompts and different voices using the Kling TTS model, which leverages advanced AI techniques to create high-quality text-to-speech.

audio
text-to-speech
MiniMaxREVIEW REQUIRED

Minimax

fal-ai/minimax/preview/speech-2.5-hd

Generate speech from text prompts and different voices using the MiniMax Speech-02 HD model, which leverages advanced AI techniques to create high-quality text-to-speech.

speech
text-to-speech
AlibabaREVIEW REQUIRED

Qwen 3 TTS - Text to Speech [0.6B]

fal-ai/qwen-3-tts/text-to-speech/0.6b

Bring speech to your texts using Qwen3-TTS Custom-Voice model with pre-trained voices or use your custom voice with Qwen3-TTS Clone Voice model

text-to-speech