EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 14 · 28 per page
image-to-image
KlingREVIEW REQUIRED

Kling O1 Image

fal-ai/kling-image/o1

Perform precise image edits using strong reference control, transforming subjects, styles, and local details while preserving visual consistency.

editrealismtypography
speech-to-speech
resemble-aiREVIEW REQUIRED

Chatterboxhd

resemble-ai/chatterboxhd/speech-to-speech

Transform voices using Resemble AI's Chatterbox. Convert audio to new voices or your own samples, with expressive results and built-in perceptual watermarking.

image-to-video
ByteDanceREVIEW REQUIRED

OmniHuman

fal-ai/bytedance/omnihuman

OmniHuman generates video using an image of a human figure paired with an audio file. It produces vivid, high-quality videos where the character’s emotions and movements maintain a strong correlation with the audio.

image-to-videolipsync
image-to-video
Luma AIREVIEW REQUIRED

Luma Ray 3.2 Image to Video

luma/agent/ray/v3.2/image-to-video

Luma Ray 3.2 animates a source image into cinematic motion guided by a text prompt, preserving the starting frame's look while controlling resolution, duration, and seamless looping.

stylizedtransformlipsync
text-to-video
ByteDanceREVIEW REQUIRED

Seedance 1.0 Pro

fal-ai/bytedance/seedance/v1/pro/text-to-video

Seedance 1.0 Pro, a high quality video generation model developed by Bytedance.

video-to-video
falREVIEW REQUIRED

Birefnet

fal-ai/birefnet/v2/video

Video background removal version of bilateral reference framework (BiRefNet) for high-resolution dichotomous image segmentation (DIS)

utilityediting
video-to-video
xAIREVIEW REQUIRED

Grok Imagine Video

xai/grok-imagine-video/edit-video

Edit videos using xAI's Grok Imagine

video-editv2vgrokxai
text-to-image
IdeogramREVIEW REQUIRED

Ideogram V2

fal-ai/ideogram/v2

Generate high-quality images, posters, and logos with Ideogram V2. Features exceptional typography handling and realistic outputs optimized for commercial and creative use.

realismtypography
video-to-video
KlingREVIEW REQUIRED

Kling O3 Reference Video to Video [Standard]

fal-ai/kling-video/o3/standard/video-to-video/reference

Kling O3 Omni generates new shots guided by an input reference video, preserving cinematic language such as motion, and camera style to produce seamless scene continuity.

video-to-video
text-to-image
RecraftREVIEW REQUIRED

Recraft V4.1 Text to Image Pro

fal-ai/recraft/v4.1/pro/text-to-image

Recraft V4.1 Pro pushes the V4.1 model into high-resolution territory — up to 2048×2048 and ultra-wide formats. Made for hero imagery, campaign work, and print, it preserves the same design taste at sizes ready for the final deliverable.

stylizedtransformtypography
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit 2511

fal-ai/qwen-image-edit-2511/lora

Endpoint for Qwen's Image Editing 2511 model with LoRa support.

stylizedtransformlora
image-to-image
AlibabaREVIEW REQUIRED

Wan v2.6 Image to Image

wan/v2.6/image-to-image

Wan 2.6 image-to-image model.

image-to-image
text-to-image
AlibabaREVIEW REQUIRED

Wan

fal-ai/wan/v2.7/text-to-image

Generate high-quality images from text prompts using the WAN 2.7 model with advanced prompt understanding and detailed output.

wantext-to-imageimage-generation
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 Kontext [max]

fal-ai/flux-pro/kontext/max/text-to-image

FLUX.1 Kontext [max] text-to-image is a new premium model brings maximum performance across all aspects – greatly improved prompt adherence.

text-to-speech
MiniMaxREVIEW REQUIRED

MiniMax Speech 2.6 [HD]

fal-ai/minimax/speech-2.6-hd

Generate speech from text prompts and different voices using the MiniMax Speech-2.6 HD model, which leverages advanced AI techniques to create high-quality text-to-speech.

text-to-speech
image-to-image
KlingREVIEW REQUIRED

Kling Image

fal-ai/kling-image/v3/image-to-image

Kling Image V3: Latest kling image model

image-to-image
text-to-video
AlibabaREVIEW REQUIRED

Wan-2.2 Text-to-Video A14B

fal-ai/wan/v2.2-a14b/text-to-video

Wan-2.2 text-to-video is a video model that generates high-quality videos with high visual quality and motion diversity from text prompts.

text to videomotion
text-to-video
ByteDanceREVIEW REQUIRED

Bytedance Seedance V1 Pro Fast Text To Video

fal-ai/bytedance/seedance/v1/pro/fast/text-to-video

Text to Video endpoint for Seedance 1.0 Pro Fast, a next-generation video model designed to deliver maximum performance at minimal cost

bytedancefastmotion
video-to-video
AlibabaREVIEW REQUIRED

Wan-2.2 Animate Move

fal-ai/wan/v2.2-14b/animate/move

Wan-Animate is a video model that generates high-fidelity character videos by replicating the expressions and movements of characters from reference videos.

video to videomotion
image-to-video
KlingREVIEW REQUIRED

Kling Video

fal-ai/kling-video/o3/4k/reference-to-video

Kling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling

stylizedtransformlipsync
vision
falREVIEW REQUIRED

Moondream 3 Preview [Query]

fal-ai/moondream3-preview/query

Moondream 3 is a vision language model that brings frontier-level visual reasoning with native object detection, pointing, and OCR capabilities to real-world applications requiring fast, inexpensive inference at scale.

Vision
video-to-video
veedREVIEW REQUIRED

Video Background Removal

veed/video-background-removal

Remove background from any video with people and objects. No green screen needed.

image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.2 [klein] 9B Base

fal-ai/flux-2/klein/9b/base/edit

Image-to-image editing with Flux 2 [klein] 9B Base from Black Forest Labs. Precise modifications using natural language descriptions and hex color control.

image-to-video
AlibabaREVIEW REQUIRED

Happy Horse

alibaba/happy-horse/reference-to-video

Generate 1080p video with synchronized native audio from a text prompt and references. Aspect ratios: 16:9, 9:16, 1:1, 4:3, 3:4. Duration: 3–15s.

stylizedtransformlipsync
text-to-speech
AlibabaREVIEW REQUIRED

Qwen 3 TTS - Voice Design [1.7B]

fal-ai/qwen-3-tts/voice-design/1.7b

Create custom voices using Qwen3-TTS Voice Design model and later use Clone Voice model to create your own voices!

text-to-speechvoice-design
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.2 [klein] 9B LoRA

fal-ai/flux-2/klein/9b/edit/lora

Image-to-image editing with FLUX.2 [klein] 9B from Black Forest Labs and custom LoRA. Precise modifications using natural language descriptions and hex color control.

video-to-video
AlibabaREVIEW REQUIRED

Wan

fal-ai/wan/v2.7/edit-video

Wan 2.7 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.

stylizedtransformlipsync
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 Krea [dev]

fal-ai/flux/krea

FLUX.1 Krea [dev] is a 12 billion parameter flow transformer that generates high-quality images from text with incredible aesthetics. It is suitable for personal and commercial use.