text-to-speechMiniMax Speech-02 HD
fal-ai/minimax/speech-02-hdGenerate speech from text prompts and different voices using the MiniMax Speech-02 HD model, which leverages advanced AI techniques to create high-quality text-to-speech.
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
text-to-speechfal-ai/minimax/speech-02-hdGenerate speech from text prompts and different voices using the MiniMax Speech-02 HD model, which leverages advanced AI techniques to create high-quality text-to-speech.
text-to-speechfal-ai/gemini-3.1-flash-ttsNewest audio model from Google introduces granular audio tags that give you precise control to direct AI speech for expressive audio generation.
fal-ai/elevenlabs/tts/turbo-v2.5Generate high-speed text-to-speech audio using ElevenLabs TTS Turbo v2.5.
text-to-speechfal-ai/minimax/speech-2.8-hdGenerate speech from text prompts and different voices using the MiniMax Speech-2.8 HD model, which leverages advanced AI techniques to create high-quality text-to-speech.
text-to-speechfal-ai/minimax/voice-cloneClone a voice from a sample audio and generate speech from text prompts using the MiniMax model, which leverages advanced AI techniques to create high-quality text-to-speech.
text-to-speechfal-ai/minimax/speech-02-turboGenerate fast speech from text prompts and different voices using the MiniMax Speech-02 Turbo model, which leverages advanced AI techniques to create high-quality text-to-speech.
text-to-speechfal-ai/qwen-3-tts/text-to-speech/1.7bBring speech to your texts using Qwen3-TTS Custom-Voice model with pre-trained voices or use your custom voice with Qwen3-TTS Clone Voice model
text-to-speechfal-ai/minimax/speech-2.8-turboGenerate speech from text prompts and different voices using the MiniMax Speech-2.8 Turbo model, which leverages advanced AI techniques to create high-quality text-to-speech.
text-to-speechxai/tts/v1Generate speech with expressive and realistic voices from xAI
text-to-speechfal-ai/inworld-ttsText to Speech Endpoint for Inworld's TTS-1.5 Max.
text-to-speechfal-ai/minimax/voice-designDesign a personalized voice from a text description, and generate speech from text prompts using the MiniMax model, which leverages advanced AI techniques to create high-quality text-to-speech.
text-to-speechfal-ai/chatterbox/text-to-speechWhether you're working on memes, videos, games, or AI agents, Chatterbox brings your content to life. Use the first tts from resemble ai.
text-to-speechfal-ai/qwen-3-tts/voice-design/1.7bCreate custom voices using Qwen3-TTS Voice Design model and later use Clone Voice model to create your own voices!
text-to-speechfal-ai/index-tts-2/text-to-speechGenerate natural, clear speeches using Index TTS 2.0 from IndexTeam
text-to-speechfal-ai/minimax/speech-2.6-hdGenerate speech from text prompts and different voices using the MiniMax Speech-2.6 HD model, which leverages advanced AI techniques to create high-quality text-to-speech.
text-to-speechfal-ai/bytedance/seed-speech/tts/v2Seed Speech developed by ByteDance, is a family of large-scale text-to-speech models capable of synthesizing speech that is virtually indistinguishable from human speech.
text-to-speechresemble-ai/chatterboxhd/text-to-speechGenerate expressive, natural speech with Resemble AI's Chatterbox. Features unique emotion control, instant voice cloning from short audio, and built-in watermarking.
text-to-speechfal-ai/minimax/speech-2.6-turboGenerate speech from text prompts and different voices using the MiniMax Speech-2.6 HD model, which leverages advanced AI techniques to create high-quality text-to-speech.
text-to-speechfal-ai/chatterbox/text-to-speech/multilingualWhether you're working on memes, videos, games, or AI agents, Chatterbox brings your content to life. Use the first tts from resemble ai.
text-to-speechalibaba/qwen-audio-3-ttsGenerate natural multilingual speech from text with fast voice and language control using Qwen Audio 3.0 TTS Flash.
text-to-speechfal-ai/zonos2Zonos2 is a text-to-speech model that clones a voice from a short sample and speaks naturally across many languages.
text-to-speechfal-ai/vibevoice/7bGenerate long, expressive multi-voice speech using Microsoft's powerful TTS
text-to-speechfal-ai/mayaMaya1 is a state-of-the-art speech model by Maya Research for expressive voice generation, built to capture real human emotion and precise voice design.
text-to-speechfal-ai/orpheus-ttsOrpheus TTS is a state-of-the-art, Llama-based Speech-LLM designed for high-quality, empathetic text-to-speech generation. This model has been finetuned to deliver human-level speech synthesis, achieving exceptional clarity, expressiveness, and real-time performances.
text-to-speechfal-ai/dia-ttsDia directly generates realistic dialogue from transcripts. Audio conditioning enables emotion control. Produces natural nonverbals like laughter and throat clearing.
text-to-speechfal-ai/kling-video/v1/ttsGenerate speech from text prompts and different voices using the Kling TTS model, which leverages advanced AI techniques to create high-quality text-to-speech.
text-to-speechfal-ai/minimax/preview/speech-2.5-hdGenerate speech from text prompts and different voices using the MiniMax Speech-02 HD model, which leverages advanced AI techniques to create high-quality text-to-speech.
text-to-speechfal-ai/qwen-3-tts/text-to-speech/0.6bBring speech to your texts using Qwen3-TTS Custom-Voice model with pre-trained voices or use your custom voice with Qwen3-TTS Clone Voice model