EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 7 · 28 per page
image-to-video
KlingREVIEW REQUIRED

Kling 2.1 (pro)

fal-ai/kling-video/v2.1/pro/image-to-video

Kling 2.1 Pro is an advanced endpoint for the Kling 2.1 model, offering professional-grade videos with enhanced visual fidelity, precise camera movements, and dynamic motion control, perfect for cinematic storytelling.

image-to-video
AlibabaREVIEW REQUIRED

Wan

fal-ai/wan/v2.2-a14b/image-to-video/turbo

Wan-2.2 Turbo image-to-video is a video model that generates high-quality videos with high visual quality and motion diversity from text prompts.

image-to-video
veedREVIEW REQUIRED

Fabric 1.0

veed/fabric-1.0

VEED Fabric 1.0 is an image-to-video API that turns any image into a talking video

lipsyncavatar
video-to-video
KlingREVIEW REQUIRED

Kling Video

fal-ai/kling-video/v3/standard/motion-control

Transfer movements from a reference video to any character image. Cost-effective mode for motion transfer, perfect for portraits and simple animations.

stylizedtransformediting
image-to-image
falREVIEW REQUIRED

Bria Eraser

fal-ai/bria/eraser

Bria Eraser enables precise removal of unwanted objects from images while maintaining high-quality outputs. Trained exclusively on licensed data for safe and risk-free commercial use. Access the model's source code and weights: https://bria.ai/contact-us

image editingobject removal
text-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 Fast Text to Video

bytedance/seedance-2.0/fast/text-to-video

ByteDance's most advanced text-to-video model, fast tier. Lower latency and cost with cinematic output, native audio, multi-shot editing, and director-level camera control.

stylizedtransformlipsync
text-to-image
IdeogramREVIEW REQUIRED

Ideogram V4.0 Text to Image

ideogram/v4

Generate high-quality images, posters, and logos with Ideogram's latest V4.0q — producing crisp visuals with accurate text rendering, fine detail, and full creative control for polished, ready-to-use designs.

realismtypographystylized
text-to-image
falREVIEW REQUIRED

Krea 2 Turbo

fal-ai/krea-2/turbo

Generate high-fidelity images from text in seconds with Krea 2 Turbo, the speed-optimized open-source version of Krea 2, preserving its aesthetic range for rapid ideation.

stylizedtransformtypography
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image 3 Image Editing

alibaba/qwen-image-3/edit

Edits images from one to three reference images and a natural-language instruction, preserving key details such as facial features and identity while applying the requested changes

stylizedtransformtypography
image-to-video
PixVerseREVIEW REQUIRED

PixVerse V6 Image To Video

fal-ai/pixverse/v6/image-to-video

Pixverse's latest V6 Model

image-to-video
image-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 Mini Image to Video

bytedance/seedance-2.0/mini/image-to-video

Seedance 2.0 Mini is a faster version of Seedance 2.0 that brings great performance and high generation speed at a lower cost.

stylizedtransformlipsync
video-to-video
ByteDanceREVIEW REQUIRED

Bytedance Upscaler Upscale Video

fal-ai/bytedance-upscaler/upscale/video

Upscale videos with Bytedance's video upscaler.

upscalervideobytedance
image-to-video
KlingREVIEW REQUIRED

Kling AI Avatar v2 Pro

fal-ai/kling-video/ai-avatar/v2/pro

Kling AI Avatar v2 Pro: The premium endpoint for creating avatar videos with realistic humans, animals, cartoons, or stylized characters

video-to-video
falREVIEW REQUIRED

Ffmpeg Api Merge Audio-Video

fal-ai/ffmpeg-api/merge-audio-video

Merge videos with standalone audio files or audio from video files.

ffmpeg
text-to-image
Black Forest LabsREVIEW REQUIRED

Flux 2 Max

fal-ai/flux-2-max

FLUX.2 [max] delivers state-of-the-art image generation and advanced image editing with exceptional realism, precision, and consistency.

flux2max
video-to-video
GoogleREVIEW REQUIRED

Gemini Omni Flash 1.1 Edit

google/gemini-omni-flash/v1.1/edit

Gemini Omni Flash 1.1 is Google's multimodal video model. This endpoint edits video through natural-language instruction, applying the requested change while preserving the parts of the scene you want kept, and carrying character and scene consistency across successive edits.

utilityeditingtransform
audio-to-audio
falREVIEW REQUIRED

Audio Understanding

fal-ai/audio-understanding

A audio understanding model to analyze audio content and answer questions about what's happening in the audio based on user prompts.

utilityaudio
image-to-video
AlibabaREVIEW REQUIRED

Wan v2.2 A14B

fal-ai/wan/v2.2-a14b/image-to-video

fal-ai/wan/v2.2-A14B/image-to-video

video-to-video
falREVIEW REQUIRED

LatentSync

fal-ai/latentsync

LatentSync is a video-to-video model that generates lip sync animations from audio using advanced algorithms for high-quality synchronization.

animationlip sync
image-to-video
falREVIEW REQUIRED

Wan 2.5 Image to Video

fal-ai/wan-25-preview/image-to-video

Wan 2.5 image-to-video model.

text-to-video
KlingREVIEW REQUIRED

Kling LipSync Audio-to-Video

fal-ai/kling-video/lipsync/audio-to-video

Kling LipSync is an audio-to-video model that generates realistic lip movements from audio input.

audio to videolipsync
text-to-video
xAIREVIEW REQUIRED

Grok Imagine Video

xai/grok-imagine-video/text-to-video

Generate videos with audio from text using Grok Imagine Video.

xaigrokt2vtext-to-video
text-to-audio
MiniMaxREVIEW REQUIRED

Minimax Music

fal-ai/minimax-music/v2

Generate music from text prompts using the MiniMax Music 2.0 model, which leverages advanced AI techniques to create high-quality, diverse musical compositions.

musicaudio
text-to-image
GoogleREVIEW REQUIRED

Gemini 3 Pro Image Preview

fal-ai/gemini-3-pro-image-preview

Gemini 3 Pro Image (a.k.a Nano Banana Pro) is Google's state-of-the-art high-fidelity image generation and editing model

realismtypography
speech-to-text
ElevenLabsREVIEW REQUIRED

ElevenLabs Speech to Text

fal-ai/elevenlabs/speech-to-text

Generate text from speech using ElevenLabs advanced speech-to-text model.

speech
image-to-3d
falREVIEW REQUIRED

Trellis

fal-ai/trellis

Generate 3D models from your images using Trellis. A native 3D generative model enabling versatile and high-quality 3D asset creation.

stylized
text-to-video
AlibabaREVIEW REQUIRED

Wan Text to Video

alibaba/wan-3.0/text-to-video

Wan 3.0 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.

stylizedtransformlipsync