EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 50 · 28 per page
video-to-video
falREVIEW REQUIRED

LTX-Video 13B 0.9.8 Distilled

fal-ai/ltxv-13b-098-distilled/extend

Extend videos using LTX Video-0.9.8 13B Distilled and custom LoRA

ltx-videoextend
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit 2509 Lora Gallery

fal-ai/qwen-image-edit-2509-lora-gallery/integrate-product

Blend products into backgrounds with automatic perspective and lighting correction

stylizedtransform
text-to-video
falREVIEW REQUIRED

Infinity Star

fal-ai/infinity-star/text-to-video

InfinityStar’s unified 8B spacetime autoregressive engine to turn any text prompt into crisp 720p videos - 10× faster than diffusion models.

text-to-video
text-to-video
falREVIEW REQUIRED

Bernini-R Text to Video

fal-ai/bernini-r/text-to-video

Generate high-quality video from a text prompt with Bernini-R, ByteDance's unified video generation and editing model.

text-to-videocinematic
training
falREVIEW REQUIRED

LTX 2.3 Trainer (V2) - Text-to-Audio

fal-ai/ltx23-trainer-v2/t2a

Train a LoRA that generates audio from a text prompt — the audio counterpart of text-to-video — learning a sound or style from your clips.

speech-to-text
falREVIEW REQUIRED

Pipecat's Smart Turn model

fal-ai/smart-turn

An open source, community-driven and native audio turn detection model by Pipecat AI.

3d-to-3d
hitem3dREVIEW REQUIRED

Hi3D 3D to Multicolor

hitem3d/hi3d/multicolor

Convert a textured 3D model into a multicolor model suited for multicolor 3D printing with Hi3D.

3d-to-3dmulticolor
text-to-video
falREVIEW REQUIRED

LTX-2 19B

fal-ai/ltx-2-19b/text-to-video/lora

Generate video with audio from text using LTX-2 and custom LoRA

image-to-image
falREVIEW REQUIRED

Florence-2 Large

fal-ai/florence-2-large/region-to-segmentation

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

multimodalvisionsegmentation
training
falREVIEW REQUIRED

LTX-2.3 22B Video to Video Trainer

fal-ai/ltx23-v2v-trainer

Train LTX-2.3 22B for video transformation or video-conditioned generation.

ltx2-videofine-tuningvideo-to-video
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit Plus Lora Gallery

fal-ai/qwen-image-edit-plus-lora-gallery/lighting-restoration

Removes harsh shadows and light spots from images, replacing them with soft, even, natural-looking illumination.

stylizedtransform
text-to-video
falREVIEW REQUIRED

Heygen

fal-ai/heygen/avatar4/digital-twin

Heygen Avatar 4 Digital Twin Model

text-to-video
training
MiniMaxREVIEW REQUIRED

MiniMax H3 First Last Frame LoRA Trainer

minimax/h3/flf2v/trainer

Train a MiniMax H3 LoRA on first/last/both keyframe signatures, teaching it to generate video with audio that starts on one image and lands on another.

utilityediting
audio-to-audio
mirelo-aiREVIEW REQUIRED

Mirelo SFX1.6

mirelo-ai/sfx1.6/extend-audio

Extend any sound effect with seamless, natural tails.

audio-to-audiosfx
training
falREVIEW REQUIRED

Stable Audio 3 Trainer

fal-ai/stable-audio-3-trainer

Stable Audio 3 LoRA Trainer fine-tunes Stable Audio 3 base models on paired audio-caption datasets, producing compact LoRA weights that adapt generation toward a custom music style, sound palette, or domain.

musicaudiosfxlora
vision
OpenAIREVIEW REQUIRED

Isaac 0.1 [OpenAI Compatible Endpoint]

perceptron/isaac-01/openai/v1/chat/completions

OpenAI spec compatible endpoint of Isaac-01 which is a multimodal vision-language model from Perceptron for various vision language tasks.

multimodalvision
text-to-video
falREVIEW REQUIRED

LTX-2.3 22B Distilled

fal-ai/ltx-2.3-22b/distilled/text-to-video/lora

Generate video with audio from text using LTX-2.3 Distilled and custom LoRA

text-to-video
mirage-apiREVIEW REQUIRED

Avatar X

mirage-api/avatar-x/text-to-video

The Avatar X API offers access to Mirage's most advanced generation model yet, delivering industry-leading identity preservation and expressivity in AI video

avatarlipsynctalking-head
image-to-image
falREVIEW REQUIRED

Invisible Watermark

fal-ai/invisible-watermark

Invisible Watermark is a model that can add an invisible watermark to an image.

utilityediting
training
falREVIEW REQUIRED

Wan-2.1 LoRA Trainer

fal-ai/wan-trainer/t2v

Train custom LoRAs for Wan-2.1 T2V 1.3B

loratraining
image-to-image
falREVIEW REQUIRED

Image Preprocessors

fal-ai/image-preprocessors/pidi

PIDI (Pidinet) preprocessor.

detectionpreprocessutilitycontrolnet
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit 2509 Lora Gallery

fal-ai/qwen-image-edit-2509-lora-gallery/face-to-full-portrait

Generate full portrait from a cropped face photo

stylizedtransform
video-to-video
falREVIEW REQUIRED

Ltx 2.3 Quality

fal-ai/ltx-2.3-quality/inpaint/lora

Inpaint high-quality video using LTX-2.3 with lora

inpaint
text-to-audio
falREVIEW REQUIRED

Ltx 2.3 Quality

fal-ai/ltx-2.3-quality/text-to-audio/lora

Text to Audio high-quality using LTX-2.3 with Lora

text-to-audio
video-to-video
falREVIEW REQUIRED

One To All Animation

fal-ai/one-to-all-animation/1.3b

One-to-All Animation is a pose driven video model that animates characters from a single reference image, enabling flexible, alignment-free motion transfer across diverse styles and scenes

video to videomotion
video-to-video
falREVIEW REQUIRED

LTX 2.3 22B

fal-ai/ltx-2.3-22b/reference-video-to-video/lora

Generate video with audio from reference video, text and images using LTX-2.3 and custom LoRA

video-to-video
falREVIEW REQUIRED

Ltx 2.3 Quality

fal-ai/ltx-2.3-quality/outpaint/lora

Outpaint high-quality video using LTX-2.3 with Lora

outpaintoutpainting