EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 40 · 28 per page
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Layered

fal-ai/qwen-image-layered/lora

Qwen-Image-Layered is a model capable of decomposing an image into multiple RGBA layers. Use loras to get your custom outputs.

qwenlora
video-to-video
decartREVIEW REQUIRED

Lucy 2.1 VTON Realtime

decart/lucy2-vton/realtime

Realtime Try On experience with Decart Lucy 2.1 VTON

text-to-video
falREVIEW REQUIRED

Infinitalk

fal-ai/infinitalk/single-text

Infinitalk model generates a talking avatar video from a text and audio file. The avatar lip-syncs to the provided audio with natural facial expressions.

image-to-image
falREVIEW REQUIRED

Optimized Latent Consistency (SDv1.5)

fal-ai/lcm-sd15-i2i

Produce high-quality images with minimal inference steps. Optimized for 512x512 input image size.

diffusionlcmreal-time
text-to-image
IdeogramREVIEW REQUIRED

Ideogram V2 Turbo

fal-ai/ideogram/v2/turbo

Accelerated image generation with Ideogram V2 Turbo. Create high-quality visuals, posters, and logos with enhanced speed while maintaining Ideogram's signature quality.

realismtypography
text-to-video
falREVIEW REQUIRED

Kandinsky5

fal-ai/kandinsky5/text-to-video/distill

Kandinsky 5.0 Distilled is a lightweight diffusion model for fast, high-quality text-to-video generation.

image-to-image
falREVIEW REQUIRED

Image Editing Wojak Style

fal-ai/image-editing/wojak-style

Transform your photos into wojak style while keeping the original characters likeness

stylizedtransform
text-to-video
falREVIEW REQUIRED

LongCat Video Distilled

fal-ai/longcat-video/distilled/text-to-video/480p

Generate long videos from text using LongCat Video Distilled

text-to-image
falREVIEW REQUIRED

Bitdance

fal-ai/bitdance

Image generation with BitDance. Fast, high-resolution photorealistic images using an autoregressive LLM— for efficient, high-quality results.

text-to-image
training
falREVIEW REQUIRED

Phota Create Profile

fal-ai/phota/create-profile

Generate profiles using 30-50 images of a subject with Phota.

stylizedtransformtypographyphota
video-to-video
falREVIEW REQUIRED

LTX Video-0.9.7 13B Distilled

fal-ai/ltx-video-13b-distilled/multiconditioning

Generate videos from prompts, images, and videos using LTX Video-0.9.7 13B Distilled and custom LoRA

videoltx-videovideo-to-videomulticondition-to-video
text-to-image
falREVIEW REQUIRED

OmniGen v1

fal-ai/omnigen-v1

OmniGen is a unified image generation model that can generate a wide range of images from multi-modal prompts. It can be used for various tasks such as Image Editing, Personalized Image Generation, Virtual Try-On, Multi Person Generation and more!

multimodaleditingtry-on
audio-to-audio
falREVIEW REQUIRED

Stable Audio 3 Small Music Base Audio Inpainting

fal-ai/stable-audio-3/small/music/base/audio-inpainting

Stable Audio 3 Small Music Base audio inpainting is the foundational 459 million parameter checkpoint for editing or filling selected music segments guided by text prompts.

musiceditingrestoration
image-to-3d
falREVIEW REQUIRED

Hunyuan3D

fal-ai/hunyuan3d/v2/multi-view/turbo

Generate 3D models from your images using Hunyuan 3D. A native 3D generative model enabling versatile and high-quality 3D asset creation.

stylized
video-to-video
Black Forest LabsREVIEW REQUIRED

Flux 3 Extend Video Draft

blackforestlabs/flux-3/extend-video/draft

FLUX.3 is Black Forest Labs' frontier audio/video model. Generate fast, low-cost draft previews that continue an existing clip, with a reusable draft cache for full-quality enhancement.

stylizedtransformlipsync
video-to-video
falREVIEW REQUIRED

Lightx

fal-ai/lightx/relight

Use tlightx capabilities to relight and recamera your videos.

video-to-video
image-to-image
falREVIEW REQUIRED

Telestyle V2 Style Transfer

fal-ai/telestyle-v2

Restyle any image with TeleStyle v2 — provide an original image and a styling reference, and the model re-renders the original in the reference's visual style while preserving its content and composition.

stylizedtransformediting
training
RecraftREVIEW REQUIRED

Recraft V4 Styles Pro Create Style

recraft/v4/pro/create-style

Creates a reusable style from your reference images for use with Recraft V4 Styles Pro generation

stylizedtransformediting
training
Black Forest LabsREVIEW REQUIRED

Flux 2 Klein 9B Base Trainer

fal-ai/flux-2-klein-9b-base-trainer/edit

Fine-tune FLUX.2 [klein] 9B from Black Forest Labs with custom datasets. Create specialized LoRA adaptations for specific editing tasks.

text-to-image
IdeogramREVIEW REQUIRED

Ideogram V2A

fal-ai/ideogram/v2a

Generate high-quality images, posters, and logos with Ideogram V2A. Features exceptional typography handling and realistic outputs optimized for commercial and creative use.

realismtypography
text-to-image
falREVIEW REQUIRED

Z-Image Turbo Seamless Tiling Lora

fal-ai/z-image/turbo/tiling/lora

Generate seamlessly tiling photorealistic images from text using Z-Image Turbo and custom LoRA

z-imageturboseamlesstiling
image-to-video
moonvalleyREVIEW REQUIRED

Marey Realism V1.5

moonvalley/marey/i2v

Generate a video starting from an image as the first frame with Marey, a generative video model trained exclusively on fully licensed data.

image-to-image
falREVIEW REQUIRED

Image Editing Professional Photo

fal-ai/image-editing/professional-photo

Turn your casual photos into stunning professional studio portraits with perfect lighting and high-end photography style.

stylizedtransform
vision
falREVIEW REQUIRED

Sam 3

fal-ai/sam-3/image/embed

SAM 3 is a unified foundation model for promptable segmentation in images and videos. It can detect, segment, and track objects using text or visual prompts such as points, boxes, and masks.

embeddingsmaskreal-time
vision
falREVIEW REQUIRED

Sa2VA 8B Image

fal-ai/sa2va/8b/image

Sa2VA is an MLLM capable of question answering, visual prompt understanding, and dense object segmentation at both image and video levels

multimodalvision
text-to-video
falREVIEW REQUIRED

LongCat Video

fal-ai/longcat-video/text-to-video/720p

Generate long videos in 720p/30fps from text using LongCat Video