EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

image to video

Page 1 · 28 per page
image-to-video
MiniMaxREVIEW REQUIRED

H3 Max Image to Video

minimax/h3-max/image-to-video

fal's H3 Max is a post-trained variant of MiniMax H3, tuned for stronger prompt adherence and better aesthetics while co-optimized with our custom inference stack for higher throughput with no compromises on output quality

stylizedtransformlipsync
image-to-video
MiniMaxREVIEW REQUIRED

H3 Max Turbo Image to Video

minimax/h3-max-turbo/image-to-video

fal's H3 Max Turbo is a post-trained variant of MiniMax H3, tuned for stronger prompt adherence and better aesthetics while co-optimized with our custom inference stack for higher throughput with no compromises on output quality

stylizedtransformlipsync
image-to-video
KlingREVIEW REQUIRED

Kling Video v3 Image to Video [Pro]

fal-ai/kling-video/v3/pro/image-to-video

Kling 3.0 Pro: Top-tier image-to-video with cinematic visuals, fluid motion, and native audio generation, with custom element support.

image-to-video
image-to-video
MiniMaxREVIEW REQUIRED

H3 Max Reference to Video

minimax/h3-max/reference-to-video

fal's H3 Max is a post-trained variant of MiniMax H3, tuned for stronger prompt adherence and better aesthetics while co-optimized with our custom inference stack for higher throughput with no compromises on output quality

stylizedtransformtypography
image-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.5 Reference to Video

bytedance/seedance-2.5/reference-to-video

Dreamina Seedance 2.5 generates video from up to 50 multimodal references images, video, audio, and style inputs, locking a character, set, and palette across a full 30-second take for production-grade consistency.

stylizedtransformlipsync
image-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.5 Image to Video

bytedance/seedance-2.5/image-to-video

Dreamina Seedance 2.5 animates a single still into a native 30-second clip at up to 720p, extending one frame into continuous, coherent motion without the drift or stitching of shorter multi-clip workflows.

stylizedtransformlipsync
image-to-video
KlingREVIEW REQUIRED

Kling Video

fal-ai/kling-video/v2.5-turbo/pro/image-to-video

Kling 2.5 Turbo Pro: Top-tier image-to-video generation with unparalleled motion fluidity, cinematic visuals, and exceptional prompt precision.

stylizedtransform
image-to-video
KlingREVIEW REQUIRED

Kling Video v3 Image to Video [Standard]

fal-ai/kling-video/v3/standard/image-to-video

Kling 3.0 Standard: Top-tier image-to-video with cinematic visuals, fluid motion, and native audio generation, with custom element support.

image-to-video
image-to-video
MiniMaxREVIEW REQUIRED

MiniMax H3 Reference to Video

minimax/h3/reference-to-video

MiniMax H3 is a frontier video model. This endpoint generates 2K video from multimodal references up to 9 images for subject and style, 3 video clips for motion, and 3 audio clips each cited in the prompt by order, keeping subjects consistent while following the referenced motion and audio.

stylizedtransformlipsync
image-to-video
ByteDanceREVIEW REQUIRED

Seedance 2 Image to Video

bytedance/seedance-2.0/image-to-video

ByteDance's most advanced image-to-video model. Animate still images into cinematic video with synchronized audio, start and end frame control, and motion prompts.

stylizedtransformlipsync
image-to-video
ByteDanceREVIEW REQUIRED

Seedance 2 Reference to Video

bytedance/seedance-2.0/reference-to-video

ByteDance's most advanced reference-to-video model. Generate video from up to 9 images, 3 videos, and 3 audio clips with native audio and cinematic camera control.

stylizedtransformlipsync
image-to-video
KlingREVIEW REQUIRED

Kling Video v2.6 Image to Video

fal-ai/kling-video/v2.6/pro/image-to-video

Kling 2.6 Pro: Top-tier image-to-video with cinematic visuals, fluid motion, and native audio generation.

image-to-video
MiniMaxREVIEW REQUIRED

MiniMax H3 Image to Video

minimax/h3/image-to-video

MiniMax H3 is a frontier video model. This endpoint animates a supplied image into 2K video, using it as the opening frame or pairs a first and last frame to control a transition between two images with the aspect ratio following the input.

stylizedtransformlipsync
image-to-video
GoogleREVIEW REQUIRED

Veo 3.1 Fast

fal-ai/veo3.1/fast/image-to-video

Generate videos from your image prompts using Veo 3.1 fast.

image-to-video
MiniMaxREVIEW REQUIRED

H3 Max Camera Controls

minimax/h3-max/camera-controls

H3 Max Multi Angle turns a single image into a video with precise, keyframe-based control over the camera's orbit, elevation, and distance in 3D space

stylizedtransformediting
image-to-video
ByteDanceREVIEW REQUIRED

Bytedance Seedance V1.5 Pro Image To Video

fal-ai/bytedance/seedance/v1.5/pro/image-to-video

Generate videos with audio with Seedance 1.5 (supports start & end frame)

bytedanceseedanceaudio
image-to-video
GoogleREVIEW REQUIRED

Veo 3.1

fal-ai/veo3.1/image-to-video

Veo 3.1 is the latest state-of-the art video generation model from Google DeepMind

image-to-video
xAIREVIEW REQUIRED

Grok Imagine Video

xai/grok-imagine-video/image-to-video

Generate videos from images with audio using xAI's Grok Imagine Video model.

grokxaiimage-to-videoi2v
image-to-video
xAIREVIEW REQUIRED

Grok Imagine Video 1.5

xai/grok-imagine-video/v1.5/image-to-video

Generate videos from images with audio using xAI's Grok Imagine 1.5 Video model.

stylizedtransformlipsync
image-to-video
GoogleREVIEW REQUIRED

Veo3.1 Lite Image to Video

fal-ai/veo3.1/lite/image-to-video

Veo 3.1 Lite balances practical utility with professional capabilities, supporting Text-to-Video and Image-to-Video

stylizedtransformlipsync
image-to-video
KlingREVIEW REQUIRED

Kling 2.1 (standard)

fal-ai/kling-video/v2.1/standard/image-to-video

Kling 2.1 Standard is a cost-efficient endpoint for the Kling 2.1 model, delivering high-quality image-to-video generation

image-to-video
GoogleREVIEW REQUIRED

Gemini Omni Flash 1.1 Image to Video

google/gemini-omni-flash/v1.1/image-to-video

Gemini Omni Flash 1.1 is Google's multimodal video model. This endpoint animates a still image into video with synchronized audio, extending a single frame into coherent motion that reflects the logic of the real world.

stylizedtransformlipsync
image-to-video
KlingREVIEW REQUIRED

Kling O3 Image to Video [Pro]

fal-ai/kling-video/o3/pro/image-to-video

Generate a video by taking a start frame and an end frame, animating the transition between them while following text-driven style and scene guidance.

image-to-video
image-to-video
AlibabaREVIEW REQUIRED

Wan 3.0

alibaba/wan-3.0/reference-to-video

Wan 3.0 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.

stylizedtransformlipsync
image-to-video
KlingREVIEW REQUIRED

Kling O3 Image to Video [Pro]

fal-ai/kling-video/o3/standard/image-to-video

Generate a video by taking a start frame and an end frame, animating the transition between them while following text-driven style and scene guidance.

image-to-video
image-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 Fast Image to Video

bytedance/seedance-2.0/fast/image-to-video

ByteDance's most advanced image-to-video model, fast tier. Lower latency and cost with synchronized audio, start and end frame control, and motion prompts.

stylizedtransformlipsync
image-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 Fast Reference to Video

bytedance/seedance-2.0/fast/reference-to-video

ByteDance's most advanced reference-to-video model, fast tier. Lower latency and cost with up to 9 images, 3 videos, and 3 audio clips as inputs.

stylizedtransformlipsync
image-to-video
KlingREVIEW REQUIRED

Kling O3 Reference to Video [Pro]

fal-ai/kling-video/o3/pro/reference-to-video

Transform images, elements, and text into consistent, high-quality video scenes, ensuring stable character identity, object details, and environments.

reference-to-video