EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 5 · 28 per page
text-to-image
OpenAIREVIEW REQUIRED

GPT-Image 1.5

fal-ai/gpt-image-1.5

GPT Image 1.5 generates high-fidelity images with strong prompt adherence, preserving composition, lighting, and fine-grained detail.

openaigpt-image
image-to-image
RecraftREVIEW REQUIRED

Recraft Crisp Upscale

fal-ai/recraft/upscale/crisp

Enhances a given raster image using 'crisp upscale' tool, boosting resolution with a focus on refining small details and faces.

upscaling
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX 2 Turbo

fal-ai/flux-2/turbo

Text-to-image generation with FLUX.2 [dev] from Black Forest Labs. Enhanced realism, crisper text generation, and native editing capabilities—all at turbo speed.

video-to-video
falREVIEW REQUIRED

Ffmpeg Api

fal-ai/ffmpeg-api/merge-videos

Use ffmpeg capabilities to merge 2 or more videos.

image-to-image
IdeogramREVIEW REQUIRED

Ideogram Remove Background

fal-ai/ideogram/remove-background

Remove backgrounds from existing images with Ideogram's remove background feature. Isolate subjects cleanly for compositing and creative reuse.

image-to-video
KlingREVIEW REQUIRED

Kling AI Avatar v2 Standard

fal-ai/kling-video/ai-avatar/v2/standard

Kling AI Avatar v2 Standard: Endpoint for creating avatar videos with realistic humans, animals, cartoons, or stylized characters

text-to-video
MiniMaxREVIEW REQUIRED

MiniMax H3 Text to Video

minimax/h3/text-to-video

MiniMax H3 is a frontier video model. This endpoint generates video from a text prompt alone, rendering at 2K in durations from 5 to 15 seconds across seven aspect ratios.

stylizedtransformlipsync
image-to-image
topazREVIEW REQUIRED

Topaz Upscale Image Precision

topaz/upscale/image/precision

Professional photo upscaling powered by Topaz Labs. Gigapixel precision models (Standard V2, High Fidelity, Low Resolution, CGI, Text Refine) enlarge images faithfully up to 4x. Best for photos that must stay true to the original.

upscaleimage
image-to-image
pixelcutREVIEW REQUIRED

Pixelcut Background Remover

pixelcut/background-removal

Pixelcut’s Background Remover enables fast, ultra high-quality removal of backgrounds from images. Perfect for e-commerce and image editing workflows. Powered by advanced AI for clean, perfect cutouts every time.

background removalutilityremove background
video-to-video
falREVIEW REQUIRED

FFmpeg API Compose

fal-ai/ffmpeg-api/compose

Compose videos from multiple media sources using FFmpeg API.

ffmpeg
image-to-video
KlingREVIEW REQUIRED

Kling Video V3 Turbo Pro Image to Video

fal-ai/kling-video/v3/turbo/pro/image-to-video

Generate high quality 1080p videos from images using Kling's Turbo 3.0 model, with improved lipsync and multishot generation capabilities.

klingv3turbo1080p
image-to-video
GoogleREVIEW REQUIRED

Gemini Omni Flash 1.1 Reference to Video

google/gemini-omni-flash/v1.1/reference-to-video

Gemini Omni Flash 1.1 is Google's multimodal video model. This endpoint generates video from combined multimodal references, images, videos and text together. Reasoning across all inputs to produce a single coherent result, with characters retaining their face, clothing, and voice throughout

stylizedtransformlipsync
image-to-video
ByteDanceREVIEW REQUIRED

Bytedance Omnihuman V1.5

fal-ai/bytedance/omnihuman/v1.5

Omnihuman v1.5 is a new and improved version of Omnihuman. It generates video using an image of a human figure paired with an audio file. It produces vivid, high-quality videos where the character’s emotions and movements maintain a strong correlation with the audio.

image-to-videolipsync
text-to-image
metaREVIEW REQUIRED

Meta Muse Image Text to Image

meta/muse-image/text-to-image

Meta's Muse Image model has faithful instruction-following and exceptional visual fidelity, with fine details like text, plots, and QR codes rendered accurately.

realismtypographystylized
text-to-image
falREVIEW REQUIRED

Stable Diffusion XL

fal-ai/fast-sdxl

Run SDXL at the speed of light

diffusionloraembeddingshigh-res
video-to-video
KlingREVIEW REQUIRED

Kling Video

fal-ai/kling-video/v3/pro/motion-control

Transfer movements from a reference video to any character image. Cost-effective mode for motion transfer, perfect for portraits and simple animations.

stylizedtransformediting
image-to-image
falREVIEW REQUIRED

Ffmpeg Api

fal-ai/ffmpeg-api/extract-frame

ffmpeg endpoint for first, middle and last frame extraction from videos

utilityediting
text-to-audio
MiniMaxREVIEW REQUIRED

Minimax Music 2.6

fal-ai/minimax-music/v2.6

MiniMax Music 2.6 creates complete tracks with singing, backing music, and detailed arrangements from lyrics and a style description.

stylizedtransformlipsync
image-to-video
AlibabaREVIEW REQUIRED

Wan 3.0 Prime

alibaba/wan-3.0-prime/image-to-video

Wan 3.0 Prime Image-to-Video turns still images into dynamic, cinematic sequences with rapid turnaround, natural motion, and excellent visual continuity. It preserves the identity, composition, and atmosphere of the source image while introducing expressive movement, camera dynamics, and richly detailed animation.

imagevideo
text-to-video
GoogleREVIEW REQUIRED

Veo 3.1

fal-ai/veo3.1

Veo 3.1 by Google, the most advanced AI video generation model in the world. With sound on!

image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit 2511 Multiple Angles

fal-ai/qwen-image-edit-2511-multiple-angles

Generates same scene from different angles (azimuth/elevation) with Qwen image Edit 2511 and the Lora Multiple Angles

stylizedtransformloramulti-angles
text-to-speech
MiniMaxREVIEW REQUIRED

MiniMax Voice Cloning

fal-ai/minimax/voice-clone

Clone a voice from a sample audio and generate speech from text prompts using the MiniMax model, which leverages advanced AI techniques to create high-quality text-to-speech.

speech
text-to-video
KlingREVIEW REQUIRED

Kling v2.5 Text to Video

fal-ai/kling-video/v2.5-turbo/pro/text-to-video

Kling 2.5 Turbo Pro: Top-tier text-to-video generation with unparalleled motion fluidity, cinematic visuals, and exceptional prompt precision.

animationstylized
text-to-image
xAIREVIEW REQUIRED

Grok Imagine Image 2.0

xai/grok-imagine-image/v2.0/text-to-image

Generate images from text using xAi's Grok Imagine 2.0 model.

text-to-imagexaigrok
image-to-image
falREVIEW REQUIRED

Bria Expand Image

fal-ai/bria/expand

Bria Expand expands images beyond their borders in high quality. Trained exclusively on licensed data for safe and risk-free commercial use. Access the model's source code and weights: https://bria.ai/contact-us

outpainting
video-to-video
KlingREVIEW REQUIRED

Kling Video v2.6 Motion Control [Standard]

fal-ai/kling-video/v2.6/standard/motion-control

Transfer movements from a reference video to any character image. Cost-effective mode for motion transfer, perfect for portraits and simple animations.

text-to-video
KlingREVIEW REQUIRED

Kling Video v3 Text to Video [Standard]

fal-ai/kling-video/v3/standard/text-to-video

Kling 3.0 Standard: Top-tier text-to-video with cinematic visuals, fluid motion, and native audio generation, with multi-shot support.

text-to-video
video-to-video
falREVIEW REQUIRED

Sync Lipsync 2.0

fal-ai/sync-lipsync/v2

Generate realistic lipsync animations from audio using advanced algorithms for high-quality synchronization with Sync Lipsync 2.0 model

animationlip sync