EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 11 · 28 per page
image-to-image
AlibabaREVIEW REQUIRED

Wan

fal-ai/wan/v2.7/edit

Transform and edit existing images with text-guided instructions using the WAN 2.7 model for creative image manipulation.

wanimage-to-imageimage-editing
video-to-video
GoogleREVIEW REQUIRED

Gemini Omni Flash

google/gemini-omni-flash/edit

Edits generated video across multiple conversational turns while preserving scene coherence. Applies iterative changes through natural-language instructions without regenerating the full sequence from scratch.

stylizedtransformlipsync
text-to-video
PixVerseREVIEW REQUIRED

PixVerse V6 Text To Video

fal-ai/pixverse/v6/text-to-video

Pixverse's latest v6 Model.

text-to-video
image-to-image
Black Forest LabsREVIEW REQUIRED

Flux Kontext Lora

fal-ai/flux-kontext-lora

Fast endpoint for the FLUX.1 Kontext [dev] model with LoRA support, enabling rapid and high-quality image editing using pre-trained LoRA adaptations for specific styles, brand identities, and product-specific outputs.

image-editingimage-to-image
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit Plus

fal-ai/qwen-image-edit-plus

Endpoint for Qwen's Image Editing Plus model also known as Qwen-Image-Edit-2509. Has superior text editing capabilities and multi-image support.

image-editingimage-to-imagehigh-quality-text
text-to-speech
falREVIEW REQUIRED

Inworld TTS-1.5 Max

fal-ai/inworld-tts

Text to Speech Endpoint for Inworld's TTS-1.5 Max.

text-to-speechinworldtts
image-to-video
Black Forest LabsREVIEW REQUIRED

Flux 3 First Last Frame to Video

blackforestlabs/flux-3/first-last-frame-to-video

FLUX 3 is Black Forest Labs' frontier video model. This endpoint generates the video between a defined start and end frame, interpolating a smooth, coherent transition from the first image to the last.

stylizedtransformlipsync
video-to-video
KlingREVIEW REQUIRED

Kling O3 Reference Video to Video [Pro]

fal-ai/kling-video/o3/pro/video-to-video/reference

Kling O3 Omni generates new shots guided by an input reference video, preserving cinematic language such as motion, and camera style to produce seamless scene continuity.

video-to-video
video-to-text
openrouterREVIEW REQUIRED

OpenRouter [Video]

openrouter/router/video

Run any video-capable LLM with fal. Analyze, summarize, and understand video files using Gemini (Google) models. Supports mp4, mpeg, mov, webm, and YouTube links. Powered by OpenRouter.

text-to-video
KlingREVIEW REQUIRED

Kling 1.6

fal-ai/kling-video/v1.6/standard/text-to-video

Generate video clips from your prompts using Kling 1.6 (std)

image-to-video
KlingREVIEW REQUIRED

Kling O1 First Frame Last Frame to Video [Pro]

fal-ai/kling-video/o1/image-to-video

Generate a video by taking a start frame and an end frame, animating the transition between them while following text-driven style and scene guidance.

image-to-video
lightricksREVIEW REQUIRED

LTX 2.5 Image to Video Fast

lightricks/ltx-2.5/image-to-video/fast

LTX-2.5 is Lightricks' open-source audio-video model. This endpoint animates a still image into video with synchronized audio in a single pass, in a speed-optimized mode for quick iteration.

stylizedtransformlip-sync
image-to-video
MiniMaxREVIEW REQUIRED

MiniMax Hailuo 2.3 Fast [Standard] (Image to Video)

fal-ai/minimax/hailuo-2.3-fast/standard/image-to-video

MiniMax Hailuo-2.3-Fast Image To Video API (Standard, 768p): Advanced fast image-to-video generation model with 768p resolution

image-to-video
video-to-video
falREVIEW REQUIRED

Depth Anything Video

fal-ai/depth-anything-video

Generates depth maps from video using Video Depth Anything (CVPR 2025). Produces per-frame depth estimation with temporal consistency across frames. Supports 3 model sizes (Small, Base, Large), 5 colormaps including grayscale, side-by-side comparison with the original video, and raw depth export as .npz. Useful for 3D reconstruction, video effects, compositing, and scene understanding.

video to videomotionedit
text-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 Mini Text to Video

bytedance/seedance-2.0/mini/text-to-video

Seedance 2.0 Mini is a faster version of Seedance 2.0 that brings great performance and high generation speed at a lower cost.

stylizedtransformlipsync
image-to-video
AlibabaREVIEW REQUIRED

Happy Horse

alibaba/happy-horse/image-to-video

Alibaba's #1-ranked Happy Horse 1.0 — generate 1080p video with synchronized native audio and multilingual lip-sync from text prompts or images.

videohappy-horse
audio-to-audio
falREVIEW REQUIRED

Sam Audio

fal-ai/sam-audio/separate

Audio separation with SAM Audio. Isolate any sound using natural language—professional-grade audio editing made simple for creators, researchers, and accessibility applications.

audio-to-audiosam-audio
image-to-3d
tripo3dREVIEW REQUIRED

Tripo H3.1 Multiview to 3D

tripo3d/h3.1/multiview-to-3d

Generate 3D models from multiple view images using Tripo H3.1.

3dmultiview-to-3d3d-generationtripo
image-to-video
AlibabaREVIEW REQUIRED

Wan v2.6 Image to Video

wan/v2.6/image-to-video

Wan 2.6 image-to-video model.

image-to-video
image-to-video
MiniMaxREVIEW REQUIRED

MiniMax Hailuo 02 [Pro] (Image to Video)

fal-ai/minimax/hailuo-02/pro/image-to-video

MiniMax Hailuo-02 Image To Video API (Pro, 1080p): Advanced image-to-video generation model with 1080p resolution

video-to-video
topazREVIEW REQUIRED

Topaz Upscale Video Generative

topaz/upscale/video/generative

Professional generative video upscaling powered by Topaz Labs. Starlight models rebuild detail that is not in the source, with Fast variants at half the price. Best for low-quality, compressed or archive footage.

upscalevideo
video-to-video
KlingREVIEW REQUIRED

Kling O1 Edit Video [Pro]

fal-ai/kling-video/o1/video-to-video/edit

Edit an existing video using natural-language instructions, transforming subjects, settings, and style while retaining the original motion structure.

text-to-speech
xAIREVIEW REQUIRED

xAI Text to Speech

xai/tts/v1

Generate speech with expressive and realistic voices from xAI

image-to-3d
MeshyREVIEW REQUIRED

V7 Multi Image to 3D

meshy/v7/multi-image-to-3d

econstructs a high-fidelity textured 3D model from multiple angle views of one object, with game-ready topology and polygon control

stylizedtransform
text-to-video
KlingREVIEW REQUIRED

Kling O3 Text to Video [Pro]

fal-ai/kling-video/o3/pro/text-to-video

Generate realistic videos using Kling O3 from Kling Team!

text-to-video
text-to-image
falREVIEW REQUIRED

Z Image Turbo Lora

fal-ai/z-image/turbo/lora

Text-to-Image endpoint with LoRA support for Z-Image Turbo, a super fast text-to-image model of 6B parameters developed by Tongyi-MAI.

z-imagelorafast
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Layered

fal-ai/qwen-image-layered

Qwen-Image-Layered is a model capable of decomposing an image into multiple RGBA layers.

qwenlayer
video-to-video
briaREVIEW REQUIRED

Bria's VRMBG 3.0

bria/video/background-removal/v3

Remove backgrounds from any video with Bria's VRMBG 3.0. Fast, accurate background removal across talking heads, podcasts, product videos, commercials, and cinematic footage.

video-to-video