EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 20 · 28 per page
video-to-video
falREVIEW REQUIRED

Flashvsr

fal-ai/flashvsr/upscale/video

Upscale your videos using FlashVSR with the fastest speeds!

upscalevideo-to-video
video-to-video
falREVIEW REQUIRED

Sam 3

fal-ai/sam-3/video-rle

SAM 3 is a unified foundation model for promptable segmentation in images and videos. It can detect, segment, and track objects using text or visual prompts such as points, boxes, and masks.

segmentationmaskreal-timerle
image-to-video
PixVerseREVIEW REQUIRED

PixVerse V5 Transition

fal-ai/pixverse/v5/transition

Create seamless transition between images using PixVerse v5

stylizedtransform
text-to-video
lightricksREVIEW REQUIRED

Ltx 2.5 Text to Video Fast

lightricks/ltx-2.5/text-to-video/fast

LTX-2.5 is Lightricks' open-source audio-video model. This endpoint generates synchronized video and audio from a text prompt in a single pass, in a speed-optimized mode built for rapid iteration and previews.

stylizedtransformlipsync
image-to-image
falREVIEW REQUIRED

Workflow Utilities Extract Nth Frame

fal-ai/workflow-utilities/extract-nth-frame

FFMPEG Untility for Extracting nth Frame

image-to-video
falREVIEW REQUIRED

LTX Video-0.9.7 13B Distilled

fal-ai/ltx-video-13b-distilled/image-to-video

Generate videos from prompts and images using LTX Video-0.9.7 13B Distilled and custom LoRA

videoltx-videoimage-to-video
video-to-video
falREVIEW REQUIRED

Heygen Lipsync - Precision

fal-ai/heygen/v3/lipsync/precision

Replace or dub audio on an existing video with high-accuracy avatar-inference lip-sync.

lipsyncstylizedtransform
text-to-speech
AlibabaREVIEW REQUIRED

Qwen Audio 3.0 TTS (Flash)

alibaba/qwen-audio-3-tts

Generate natural multilingual speech from text with fast voice and language control using Qwen Audio 3.0 TTS Flash.

text-to-speechaudiospeech-synthesismultilingual
image-to-video
MiniMaxREVIEW REQUIRED

MiniMax (Hailuo AI) Video 01

fal-ai/minimax/video-01/image-to-video

Generate video clips from your images using MiniMax Video model

motiontransformation
text-to-speech
falREVIEW REQUIRED

Zonos2 Text to Speech

fal-ai/zonos2

Zonos2 is a text-to-speech model that clones a voice from a short sample and speaks naturally across many languages.

text-to-speechttsvoice cloning
image-to-image
IdeogramREVIEW REQUIRED

Ideogram V3 Character Edit

fal-ai/ideogram/character/edit

Modify consistent characters while preserving their core identity. Edit poses, expressions, or clothing without losing recognizable character features

character-consistency
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev] Control LoRA Depth

fal-ai/flux-control-lora-depth

FLUX Control LoRA Depth is a high-performance endpoint that uses a control image to transfer structure to the generated image, using a depth map.

lorastyle transfer
text-to-video
falREVIEW REQUIRED

Pika Text to Video (v2.2)

fal-ai/pika/v2.2/text-to-video

Start with a simple text input to create dynamic generations that defy expectations in up to 1080p. Experience better image clarity and crisper, sharper visuals.

editingeffectsanimation
image-to-image
falREVIEW REQUIRED

Stable Diffusion XL

fal-ai/fast-sdxl/image-to-image

Run SDXL at the speed of light

diffusionhigh-resloraip-adapter
text-to-3d
falREVIEW REQUIRED

Hunyuan Motion [1B]

fal-ai/hunyuan-motion

Generate 3D human motions via text-to-generation interface of Hunyuan Motion!

text-to-3dmotion
video-to-video
GoogleREVIEW REQUIRED

Veo 3.1 Fast

fal-ai/veo3.1/fast/extend-video

Extend Veo-Created Videos up to 30 seconds

extend-video
image-to-image
falREVIEW REQUIRED

Florence-2 Large

fal-ai/florence-2-large/open-vocabulary-detection

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

multimodalvisiondetection
video-to-video
Luma AIREVIEW REQUIRED

Luma Ray 3.2 Reframe

luma/agent/ray/v3.2/reframe

Luma Ray 3.2 reframes an existing video into a new aspect ratio guided by a text prompt, preserving the original footage frame-for-frame while controlling resolution and outpainting the surrounding canvas.

stylizedtransformlipsync
image-to-image
falREVIEW REQUIRED

Creative Upscaler

fal-ai/creative-upscaler

Create creative upscaled images.

upscaling
video-to-video
soniloREVIEW REQUIRED

V1.1 Video to Video Music

sonilo/v1.1/video-to-video-music

Generates perfectly synced music for any video. Return a licensed music soundtrack ready for commercial use (optional preservation of the original speech in video)

musiceditingrestoration
video-to-video
Luma AIREVIEW REQUIRED

Luma Ray 3.2 Video to Video

luma/agent/ray/v3.2/video-to-video

Luma Ray 3.2 re-renders an existing video into new cinematic motion guided by a text prompt, preserving the source's look and movement while controlling resolution, duration, and HDR.

stylizedtransformlipsync
image-to-image
falREVIEW REQUIRED

PATINA

fal-ai/patina/material/extract

Extract seamless tiling textures with PBR attribute maps from images

materialpbrextraction
video-to-video
falREVIEW REQUIRED

Void Video Inpainting

fal-ai/void-video-inpainting

VOID removes objects from videos along with all interactions they induce on the scene

utilityediting
text-to-image
microsoftREVIEW REQUIRED

MAI Image 2.5 Pro (Text to Image)

microsoft/mai-image-2.5-pro

Generate high-fidelity, design-ready images with precise typography, strong prompt alignment, and rich visual detail using Microsoft's flagship MAI Image 2.5 Pro.

photorealismtypographyillustrationcommercial
text-to-image
falREVIEW REQUIRED

Hidream I1 Fast

fal-ai/hidream-i1-fast

HiDream-I1 fast is a new open-source image generative foundation model with 17B parameters that achieves state-of-the-art image generation quality within 16 steps.

image-to-image
falREVIEW REQUIRED

DDColor

fal-ai/ddcolor

Bring colors into old or new black and white photos with DDColor.

image-recolorizationfacesutility
audio-to-audio
falREVIEW REQUIRED

Workflow Utilities Audio Compressor

fal-ai/workflow-utilities/audio-compressor

FFMPEG Utility for Audio Compression

text-to-video
AlibabaREVIEW REQUIRED

Happy Horse 1.1 Text to Video

alibaba/happy-horse/v1.1/text-to-video

Happy Horse 1.1 is Alibaba's #1-ranked video model. This text-to-video endpoint generates 1080p video with synchronized native audio and multilingual lip-sync from a text prompt alone.

happy-horsevideotext