EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 4 · 28 per page
text-to-video
KlingREVIEW REQUIRED

Kling Video v3 Text to Video [Pro]

fal-ai/kling-video/v3/pro/text-to-video

Kling 3.0 Pro: Top-tier text-to-video with cinematic visuals, fluid motion, and native audio generation, with multi-shot support.

text-to-video
video-to-video
AlibabaREVIEW REQUIRED

Wan 3.0 Prime

alibaba/wan-3.0-prime/reference-to-video

Wan 3.0 Prime Reference-to-Video combines reference images, videos, and audio into a unified video with fast generation and strong multimodal coherence. It follows character identity, visual style, movement, and sound cues across references to create controlled, consistent, and production-ready results.

referencevideo
text-to-image
IdeogramREVIEW REQUIRED

Ideogram Text to Image

fal-ai/ideogram/v3

Generate high-quality images, posters, and logos with Ideogram V3. Features exceptional typography handling and realistic outputs optimized for commercial and creative use.

realismtypography
text-to-speech
MiniMaxREVIEW REQUIRED

MiniMax Speech 2.8 [HD]

fal-ai/minimax/speech-2.8-hd

Generate speech from text prompts and different voices using the MiniMax Speech-2.8 HD model, which leverages advanced AI techniques to create high-quality text-to-speech.

text-to-image
kreaREVIEW REQUIRED

Krea 2 Large

krea/v2/large/text-to-image

Generate high-fidelity images from text with Krea 2 Large, supporting aspect ratio, creativity, seed controls, and optional style references.

text-to-imageimage-generationstyle-referencekrea
image-to-video
GoogleREVIEW REQUIRED

Gemini Omni Flash 1.1 Image to Video

google/gemini-omni-flash/v1.1/image-to-video

Gemini Omni Flash 1.1 is Google's multimodal video model. This endpoint animates a still image into video with synchronized audio, extending a single frame into coherent motion that reflects the logic of the real world.

stylizedtransformlipsync
image-to-video
KlingREVIEW REQUIRED

Kling O3 Image to Video [Pro]

fal-ai/kling-video/o3/pro/image-to-video

Generate a video by taking a start frame and an end frame, animating the transition between them while following text-driven style and scene guidance.

image-to-video
image-to-video
AlibabaREVIEW REQUIRED

Wan 3.0

alibaba/wan-3.0/reference-to-video

Wan 3.0 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.

stylizedtransformlipsync
text-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 Text to Video API

bytedance/seedance-2.0/text-to-video

ByteDance's most advanced text-to-video model. Cinematic output with native audio, multi-shot editing, real-world physics, and director-level camera control.

stylizedtransformlipsync
image-to-image
falREVIEW REQUIRED

Remove Background

fal-ai/imageutils/rembg

Remove the background from an image.

background removalutilityediting
image-to-video
KlingREVIEW REQUIRED

Kling O3 Image to Video [Pro]

fal-ai/kling-video/o3/standard/image-to-video

Generate a video by taking a start frame and an end frame, animating the transition between them while following text-driven style and scene guidance.

image-to-video
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX 2 Edit

fal-ai/flux-2/edit

Image-to-image editing with FLUX.2 [dev] from Black Forest Labs. Precise modifications using natural language descriptions and hex color control.

video-to-video
KlingREVIEW REQUIRED

Kling O3 Edit Video [Pro]

fal-ai/kling-video/o3/pro/video-to-video/edit

Edit videos using Kling O3 from Kling Team!

video-to-video
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [pro] Fill

fal-ai/flux-pro/v1/fill

FLUX.1 [pro] Fill is a high-performance endpoint for the FLUX.1 [pro] model that enables rapid transformation of existing images, delivering high-quality style transfers and image modifications with the core FLUX capabilities.

editing
image-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 Fast Image to Video

bytedance/seedance-2.0/fast/image-to-video

ByteDance's most advanced image-to-video model, fast tier. Lower latency and cost with synchronized audio, start and end frame control, and motion prompts.

stylizedtransformlipsync
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev]

fal-ai/flux/dev/image-to-image

FLUX.1 Image-to-Image is a high-performance endpoint for the FLUX.1 [dev] model that enables rapid transformation of existing images, delivering high-quality style transfers and image modifications with the core FLUX capabilities.

style transfer
video-to-video
falREVIEW REQUIRED

sync-3 Lipsync

fal-ai/sync-lipsync/v3

sync-3 most powerful lipsync model yet, featuring native visual intelligence for professional-quality video.

stylizedtransformlipsync
image-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 Fast Reference to Video

bytedance/seedance-2.0/fast/reference-to-video

ByteDance's most advanced reference-to-video model, fast tier. Lower latency and cost with up to 9 images, 3 videos, and 3 audio clips as inputs.

stylizedtransformlipsync
image-to-video
KlingREVIEW REQUIRED

Kling O3 Reference to Video [Pro]

fal-ai/kling-video/o3/pro/reference-to-video

Transform images, elements, and text into consistent, high-quality video scenes, ensuring stable character identity, object details, and environments.

reference-to-video
image-to-video
AlibabaREVIEW REQUIRED

Wan 3.0

alibaba/wan-3.0/image-to-video

Wan 3.0 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.

stylizedtransformlipsync
image-to-video
ByteDanceREVIEW REQUIRED

Seedance 1.0 Pro

fal-ai/bytedance/seedance/v1/pro/image-to-video

Seedance 1.0 Pro, a high quality video generation model developed by Bytedance.

text-to-audio
falREVIEW REQUIRED

Stable Audio 2.5

fal-ai/stable-audio-25/text-to-audio

Generate high quality music and sound effects using Stable Audio 2.5 from StabilityAI

audio
image-to-video
Black Forest LabsREVIEW REQUIRED

FLUX 3 Image to Video

blackforestlabs/flux-3/image-to-video

FLUX 3 is Black Forest Labs' frontier video model. This endpoint animates a single still image into video, extending one frame into coherent, natural motion.

stylizedtransformlipsync
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 Kontext [max]

fal-ai/flux-pro/kontext/max

FLUX.1 Kontext [max] is a model with greatly improved prompt adherence and typography generation meet premium consistency for editing without compromise on speed.

text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX 2 Flash

fal-ai/flux-2/flash

Text-to-image generation with FLUX.2 [dev] from Black Forest Labs. Enhanced realism, crisper text generation, and native editing capabilities— in a flash.

text-to-video
GoogleREVIEW REQUIRED

Veo 3.1 Fast

fal-ai/veo3.1/fast

Faster and more cost effective version of Google's Veo 3.1!

text-to-image
RecraftREVIEW REQUIRED

Recraft V3

fal-ai/recraft/v3/text-to-image

Recraft V3 is a text-to-image model with the ability to generate long texts, vector art, images in brand style, and much more. As of today, it is SOTA in image generation, proven by Hugging Face's industry-leading Text-to-Image Benchmark by Artificial Analysis.

vectortypographystyle