EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 10 · 28 per page
image-to-image
Black Forest LabsREVIEW REQUIRED

Flux 2 Flex

fal-ai/flux-2-flex/edit

Image editing with FLUX.2 [flex] from Black Forest Labs. Supports multi-reference editing with customizable inference steps and enhanced text rendering.

vision
falREVIEW REQUIRED

Video Understanding

fal-ai/video-understanding

A video understanding model to analyze video content and answer questions about what's happening in the video based on user prompts.

utilityvision
image-to-video
falREVIEW REQUIRED

sync-3 Avatar Image to Video

fal-ai/sync-lipsync/v3/image-to-video

sync-3 image to video turns a single still into a talking character, and works with any illustration or animated frame paired with a voice track

animationlip synctext-to-speech
image-to-image
KlingREVIEW REQUIRED

Kling Image

fal-ai/kling-image/o3/image-to-image

Kling Omni 3: Top-tier image-to-image with flawless consistency.

image-to-image
image-to-video
xAIREVIEW REQUIRED

Grok Imagine Video 1.5 Reference to Video

xai/grok-imagine-video/v1.5/reference-to-video

Generate videos from images and audio references using xAI's Grok Imagine 1.5 Video model.

stylizedtransformlipsync
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX 2 Lora

fal-ai/flux-2/lora

Text-to-image generation with LoRA support for FLUX.2 [dev] from Black Forest Labs. Custom style adaptation and fine-tuned model variations.

text-to-image
OpenAIREVIEW REQUIRED

gpt-image-1

fal-ai/gpt-image-1/text-to-image

OpenAI's latest image generation and editing model: gpt-1-image.

text-to-image
falREVIEW REQUIRED

Krea 2 Text to Image Turbo LoRA

fal-ai/krea-2/turbo/lora

Generate high-fidelity images from text with Krea 2 using a custom-trained LoRA. Apply your LoRA weights to carry a learned subject, character, or style into new generations, with aspect ratio, creativity, and seed controls.

stylizedtransformtypographyrealism
video-to-video
falREVIEW REQUIRED

sync.so -- lipsync 1.9.0-beta

fal-ai/sync-lipsync

Generate realistic lipsync animations from audio using advanced algorithms for high-quality synchronization.

animationlip sync
text-to-audio
falREVIEW REQUIRED

Kokoro TTS

fal-ai/kokoro/american-english

Kokoro is a lightweight text-to-speech model that delivers comparable quality to larger models while being significantly faster and more cost-efficient.

speech
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 Kontext [pro]

fal-ai/flux-pro/kontext/multi

Experimental version of FLUX.1 Kontext [pro] with multi image handling capabilities

text-to-speech
AlibabaREVIEW REQUIRED

Qwen 3 TTS - Text to Speech [1.7B]

fal-ai/qwen-3-tts/text-to-speech/1.7b

Bring speech to your texts using Qwen3-TTS Custom-Voice model with pre-trained voices or use your custom voice with Qwen3-TTS Clone Voice model

text-to-speech
text-to-audio
ElevenLabsREVIEW REQUIRED

Elevenlabs Music v2.5

elevenlabs/music/v2.5

Generate high quality, realistic music with fine controls using Elevenlabs Music v2.5!

musictext-to-music
image-to-video
KlingREVIEW REQUIRED

Kling 1.6

fal-ai/kling-video/v1.6/pro/image-to-video

Generate video clips from your images using Kling 1.6 (pro)

image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev] with LoRAs

fal-ai/flux-lora/image-to-image

FLUX LoRA Image-to-Image is a high-performance endpoint that transforms existing images using FLUX models, leveraging LoRA adaptations to enable rapid and precise image style transfer, modifications, and artistic variations.

lorastyle transfer
text-to-audio
soniloREVIEW REQUIRED

Sonilo V1.1 Text to Music

sonilo/v1.1/text-to-music

Generates licensed, commercial-use-safe music from a single text prompt, with full control over style, mood, instrumentation, and exact duration.

stylizedtransformlipsync
text-to-image
GoogleREVIEW REQUIRED

Gemini 3.1 Flash Image Preview

fal-ai/gemini-3.1-flash-image-preview

Gemini 3.1 Flash Image (a.k.a Nano Banana 2) is Google's new state-of-the-art fast image generation and editing model

vision
falREVIEW REQUIRED

NSFW Checker

fal-ai/x-ailab/nsfw

Predict whether an image is NSFW or SFW.

filtersafetyutility
image-to-video
falREVIEW REQUIRED

Wan-2.1 Image-to-Video

fal-ai/wan-i2v

Wan-2.1 is a image-to-video model that generates high-quality videos with high visual quality and motion diversity from images

image to videomotion
video-to-video
AlibabaREVIEW REQUIRED

Wan-2.2 Animate Replace

fal-ai/wan/v2.2-14b/animate/replace

Wan-Animate Replace is a model that can integrate animated characters into reference videos, replacing the original character while preserving the scene’s lighting and color tone for seamless environmental integration.

video to videomotion
text-to-audio
soniloREVIEW REQUIRED

V1.1 Text to Sound Effects

sonilo/v1.1/text-to-sound-effects

Generates high-quality, commercial-use-safe sound effects from a text prompt, with full control over type, texture, intensity, and exact duration.

sfxaudioeffects
audio-to-audio
falREVIEW REQUIRED

FFmpeg API [Merge Audios]

fal-ai/ffmpeg-api/merge-audios

Merge audios into a single audio using FFmpeg API!

ffmpeg
image-to-image
OpenAIREVIEW REQUIRED

gpt-image-1

fal-ai/gpt-image-1/edit-image

OpenAI's latest image generation and editing model: gpt-1-image.

image-to-image
falREVIEW REQUIRED

Sam 3 1

fal-ai/sam-3-1/image

SAM 3.1 builds comes with Object Multiplex, a shared-memory approach for joint multi-object tracking that delivers faster speeds with larger number of objects tracked.

segmentationmaskreal-time
text-to-speech
MiniMaxREVIEW REQUIRED

MiniMax Speech 2.8 [Turbo]

fal-ai/minimax/speech-2.8-turbo

Generate speech from text prompts and different voices using the MiniMax Speech-2.8 Turbo model, which leverages advanced AI techniques to create high-quality text-to-speech.

training
falREVIEW REQUIRED

Krea 2 Trainer

fal-ai/krea-2-trainer

Train a custom LoRA on your own images to teach Krea 2 a new subject, character, or style. Provide a set of training images (and an optional trigger word), and the trainer outputs LoRA weights you can use for inference with the Krea 2 LoRA endpoint.

lorapersonalization
image-to-image
falREVIEW REQUIRED

Z Image Turbo Image To Image

fal-ai/z-image/turbo/image-to-image

Generate images from text and images using Z-Image Turbo, Tongyi-MAI's super-fast 6B model.

turboz-imagefast
text-to-audio
ElevenLabsREVIEW REQUIRED

Elevenlabs

fal-ai/elevenlabs/text-to-dialogue/eleven-v3

Generate realistic audio dialogues using Eleven-v3 from ElevenLabs.

audio