EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

text to video

Page 1 · 28 per page
text-to-video
MiniMaxREVIEW REQUIRED

MiniMax H3 Max Text to Video

minimax/h3-max/text-to-video

fal's H3 Max is a post-trained variant of MiniMax H3, tuned for stronger prompt adherence and better aesthetics while co-optimized with our custom inference stack for higher throughput with no compromises on output quality

stylizedtransformlipsync
text-to-video
MiniMaxREVIEW REQUIRED

H3 Max Turbo Text to Video

minimax/h3-max-turbo/text-to-video

fal's H3 Max Turbo is a post-trained variant of MiniMax H3, tuned for stronger prompt adherence and better aesthetics while co-optimized with our custom inference stack for higher throughput with no compromises on output quality

stylizedtransformlipsync
text-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.5 Text to Video

bytedance/seedance-2.5/text-to-video

Dreamina Seedance 2.5 generates native 30-second single-shot video at up to 720p from a single text prompt, reasoning about the whole shot at once so motion, lighting, and subject identity stay coherent from first frame to last.

stylizedtransformlipsync
text-to-video
KlingREVIEW REQUIRED

Kling Video v3 Text to Video [Pro]

fal-ai/kling-video/v3/pro/text-to-video

Kling 3.0 Pro: Top-tier text-to-video with cinematic visuals, fluid motion, and native audio generation, with multi-shot support.

text-to-video
text-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 Text to Video API

bytedance/seedance-2.0/text-to-video

ByteDance's most advanced text-to-video model. Cinematic output with native audio, multi-shot editing, real-world physics, and director-level camera control.

stylizedtransformlipsync
text-to-video
GoogleREVIEW REQUIRED

Veo 3.1 Fast

fal-ai/veo3.1/fast

Faster and more cost effective version of Google's Veo 3.1!

text-to-video
MiniMaxREVIEW REQUIRED

MiniMax H3 Text to Video

minimax/h3/text-to-video

MiniMax H3 is a frontier video model. This endpoint generates video from a text prompt alone, rendering at 2K in durations from 5 to 15 seconds across seven aspect ratios.

stylizedtransformlipsync
text-to-video
GoogleREVIEW REQUIRED

Veo 3.1

fal-ai/veo3.1

Veo 3.1 by Google, the most advanced AI video generation model in the world. With sound on!

text-to-video
KlingREVIEW REQUIRED

Kling v2.5 Text to Video

fal-ai/kling-video/v2.5-turbo/pro/text-to-video

Kling 2.5 Turbo Pro: Top-tier text-to-video generation with unparalleled motion fluidity, cinematic visuals, and exceptional prompt precision.

animationstylized
text-to-video
KlingREVIEW REQUIRED

Kling Video v3 Text to Video [Standard]

fal-ai/kling-video/v3/standard/text-to-video

Kling 3.0 Standard: Top-tier text-to-video with cinematic visuals, fluid motion, and native audio generation, with multi-shot support.

text-to-video
text-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 Fast Text to Video

bytedance/seedance-2.0/fast/text-to-video

ByteDance's most advanced text-to-video model, fast tier. Lower latency and cost with cinematic output, native audio, multi-shot editing, and director-level camera control.

stylizedtransformlipsync
text-to-video
KlingREVIEW REQUIRED

Kling LipSync Audio-to-Video

fal-ai/kling-video/lipsync/audio-to-video

Kling LipSync is an audio-to-video model that generates realistic lip movements from audio input.

audio to videolipsync
text-to-video
xAIREVIEW REQUIRED

Grok Imagine Video

xai/grok-imagine-video/text-to-video

Generate videos with audio from text using Grok Imagine Video.

xaigrokt2vtext-to-video
text-to-video
AlibabaREVIEW REQUIRED

Wan Text to Video

alibaba/wan-3.0/text-to-video

Wan 3.0 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.

stylizedtransformlipsync
text-to-video
GoogleREVIEW REQUIRED

Gemini Omni Flash 1.1 Text to Video

google/gemini-omni-flash/v1.1/text-to-video

Gemini Omni Flash 1.1 is Google's multimodal video model. This endpoint generates video with synchronized native audio from a text prompt, grounded in Gemini's real-world knowledge and physics understanding, with cinematic camera control expressed in natural language.

stylizedtransformlipsync
text-to-video
Black Forest LabsREVIEW REQUIRED

Flux 3 Text to Video

blackforestlabs/flux-3/text-to-video

FLUX 3 is Black Forest Labs' frontier video model. This endpoint generates video directly from a text prompt, translating a written description into motion, composition, and scene.

stylizedtransformlipsync
text-to-video
GoogleREVIEW REQUIRED

Veo3.1 Lite Text to Video

fal-ai/veo3.1/lite

Veo 3.1 Lite balances practical utility with professional capabilities, supporting Text-to-Video and Image-to-Video

stylizedtransformlipsync
text-to-video
ByteDanceREVIEW REQUIRED

Bytedance Seedance V1.5 Pro Text To Video

fal-ai/bytedance/seedance/v1.5/pro/text-to-video

Generate videos with audio with Seedance 1.5

bytedanceseedanceaudio
text-to-video
GoogleREVIEW REQUIRED

Gemini Omni Flash

google/gemini-omni-flash

Creates video with synchronized audio from text input. Grounded in Gemini's real-world knowledge, with improved physics understanding for more coherent motion and interaction.

stylizedtransformlipsync
text-to-video
KlingREVIEW REQUIRED

Kling Video v2.6 Text to Video

fal-ai/kling-video/v2.6/pro/text-to-video

Kling 2.6 Pro: Top-tier text-to-video with cinematic visuals, fluid motion, and native audio generation.

text-to-video
AlibabaREVIEW REQUIRED

Wan 3.0 Prime

alibaba/wan-3.0-prime/text-to-video

Wan 3.0 Prime Text-to-Video transforms written prompts into polished videos with accelerated generation, fluid motion, strong scene fidelity, and coherent visual storytelling. Built for fast creative iteration, it brings complex ideas to life while preserving visual detail and cinematic consistency throughout each shot.

textvideo
text-to-video
KlingREVIEW REQUIRED

Kling 1.6

fal-ai/kling-video/v1.6/standard/text-to-video

Generate video clips from your prompts using Kling 1.6 (std)

text-to-video
PixVerseREVIEW REQUIRED

PixVerse V6 Text To Video

fal-ai/pixverse/v6/text-to-video

Pixverse's latest v6 Model.

text-to-video
text-to-video
KlingREVIEW REQUIRED

Kling O3 Text to Video [Pro]

fal-ai/kling-video/o3/pro/text-to-video

Generate realistic videos using Kling O3 from Kling Team!

text-to-video
text-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 Mini Text to Video

bytedance/seedance-2.0/mini/text-to-video

Seedance 2.0 Mini is a faster version of Seedance 2.0 that brings great performance and high generation speed at a lower cost.

stylizedtransformlipsync
text-to-video
AlibabaREVIEW REQUIRED

Wan Text to Video

fal-ai/wan/v2.7/text-to-video

Wan 2.7 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.

stylizedtransformlipsync
text-to-video
KlingREVIEW REQUIRED

Kling O3 Text to Video [Standard]

fal-ai/kling-video/o3/standard/text-to-video

Generate realistic videos using Kling O3 from Kling Team!

text-to-video
text-to-video
KlingREVIEW REQUIRED

Kling Video V3 Standard Turbo Text to Video

fal-ai/kling-video/v3/turbo/standard/text-to-video

Kling 3.0 Turbo Standard is a fast, cost-efficient video generation model that turns text prompts directly into 720P video with native audio, optimized for rapid iteration and high-volume production

stylizedtransformlipsync