EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 16 · 28 per page
image-to-image
falREVIEW REQUIRED

IC-Light-v2 for Image Relighting

fal-ai/iclight-v2

An endpoint for re-lighting photos and changing their backgrounds per a given description

relightingediting
text-to-video
AlibabaREVIEW REQUIRED

Happy Horse

alibaba/happy-horse/text-to-video

Generate 1080p video with synchronized native audio from a text prompt. Aspect ratios: 16:9, 9:16, 1:1, 4:3, 3:4. Duration: 3–15s.

happy-horse
image-to-image
IdeogramREVIEW REQUIRED

Ideogram V4.0q Image to Image

ideogram/v4/image-to-image

Ideogram V4.0q Image-to-Image transforms an input image with a text prompt, restyling and reworking the composition while preserving its core structure for prompt-faithful, high-fidelity edits.

realismtypographystylized
video-to-video
KlingREVIEW REQUIRED

Kling Video 4K Video to Video Edit

fal-ai/kling-video/o3/4k/video-to-video/edit

Kling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling

utilityediting
video-to-audio
soniloREVIEW REQUIRED

V1.1 Video to Sound Effects

sonilo/v1.1/video-to-sound-effects

Analyzes a video and generates synchronized, royalty-free sound effects timed to visible actions. Returns the generated sound-effects audio track for commercial use.

sfxaudioeffects
video-to-video
ByteDanceREVIEW REQUIRED

Bytedance Dreamactor V2

fal-ai/bytedance/dreamactor/v2

Transfer motion from a video to characters in an image using Dreamactor v2. Great performance for non-human and multiple characters

motion-controldreamactor
text-to-video
AlibabaREVIEW REQUIRED

Wan

fal-ai/wan/v2.2-a14b/text-to-video/turbo

Wan-2.2 turbo text-to-video is a video model that generates high-quality videos with high visual quality and motion diversity from text prompts.

text to videomotion
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.2 [klein] 9B Base LoRA

fal-ai/flux-2/klein/9b/base/edit/lora

Image-to-image editing with LoRA support for FLUX.2 [klein] 9B Base from Black Forest Labs. Specialized style transfer and domain-specific modifications.

image-to-3d
falREVIEW REQUIRED

Hunyuan 3D Rapid Image to 3D

fal-ai/hunyuan-3d/v3.1/rapid/image-to-3d

Rapidly generate 3D models from images using Hunyuan 3D.

3dhunyuanimage-to-3d
image-to-3d
falREVIEW REQUIRED

Hunyuan3D

fal-ai/hunyuan3d/v2

Generate 3D models from your images using Hunyuan 3D. A native 3D generative model enabling versatile and high-quality 3D asset creation.

stylized
image-to-video
Black Forest LabsREVIEW REQUIRED

Flux 3 Image To Video Draft

blackforestlabs/flux-3/image-to-video/draft

FLUX.3 is Black Forest Labs' frontier audio/video model. Generate fast, low-cost draft previews that animate a still image, with a reusable draft cache for full-quality enhancement.

stylizedtransformlipsync
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev] Fill with LoRAs

fal-ai/flux-lora-fill

FLUX.1 [dev] Fill is a high-performance endpoint for the FLUX.1 [pro] model that enables rapid transformation of existing images, delivering high-quality style transfers and image modifications with the core FLUX capabilities.

editinglora
image-to-image
falREVIEW REQUIRED

Bria Background Replace

fal-ai/bria/background/replace

Bria Background Replace allows for efficient swapping of backgrounds in images via text prompts or reference image, delivering realistic and polished results. Trained exclusively on licensed data for safe and risk-free commercial use

image editing
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit

fal-ai/qwen-image-edit/inpaint

Inpainting Endpoint for the Qwen Edit Image editing model.

image-to-imageinpaintingqwen-image
image-to-image
IdeogramREVIEW REQUIRED

Ideogram

fal-ai/ideogram/v3/reframe

Extend existing images with Ideogram V3's reframe feature. Create expanded versions and adaptations while preserving main image and adding new creative directions through prompt guidance.

realismtypography
image-to-3d
falREVIEW REQUIRED

Sam 3

fal-ai/sam-3/3d-body

SAM 3D allows for accurate 3D reconstruction of human body shape and position from a single image.

3dhumanpose
vision
falREVIEW REQUIRED

Moondream3 Preview [Detect]

fal-ai/moondream3-preview/detect

Moondream 3 is a vision language model that brings frontier-level visual reasoning with native object detection, pointing, and OCR capabilities to real-world applications requiring fast, inexpensive inference at scale.

Vision
image-to-video
MiniMaxREVIEW REQUIRED

Minimax

fal-ai/minimax/hailuo-02-fast/image-to-video

Create blazing fast and economical videos with MiniMax Hailuo-02 Image To Video API at 512p resolution

stylizedtransform
text-to-image
KlingREVIEW REQUIRED

Kling Image

fal-ai/kling-image/v3/text-to-image

Kling V3: Latest Kling Image model

text-to-image
text-to-image
falREVIEW REQUIRED

Hunyuan Image

fal-ai/hunyuan-image/v3/text-to-image

Leverage the state-of-the-art capabilities of Hunyuan Image 3.0 to generate visual content that effectively conveys the messaging of your written material.

text-to-image
text-to-image
IdeogramREVIEW REQUIRED

V4.0q [instant]

ideogram/v4/instant

Generate high-quality images, posters, and logos with Ideogram's latest V4.0q — producing crisp visuals with accurate text rendering, fine detail, and full creative control for polished, ready-to-use designs FRACTION OF A SECOND.

realismtypographystylized
image-to-image
falREVIEW REQUIRED

Florence-2 Large

fal-ai/florence-2-large/caption-to-phrase-grounding

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

multimodalvision
image-to-3d
tripo3dREVIEW REQUIRED

Triposplat

tripo3d/triposplat

TripoSplat is an open-source model from TripoAI / VAST AI Research that converts a single 2D image into high-quality 3D Gaussians using a novel learned density-control approach

3Dgaussian-splat
image-to-image
falREVIEW REQUIRED

Segment Anything Model 2

fal-ai/sam2/auto-segment

SAM 2 is a model for segmenting images automatically. It can return individual masks or a single mask for the entire image.

segmentationmask
image-to-video
KlingREVIEW REQUIRED

Kling O1 First Frame Last Frame to Video [Standard]

fal-ai/kling-video/o1/standard/image-to-video

Generate a video by taking a start frame and an end frame, animating the transition between them while following text-driven style and scene guidance.

audio-to-audio
KlingREVIEW REQUIRED

Kling Video Create Voice

fal-ai/kling-video/create-voice

Create Voices to be used with Kling Models Voice Control

text-to-audio
falREVIEW REQUIRED

Stable Audio 3 Small SFX Text to Audio

fal-ai/stable-audio-3/small/sfx/text-to-audio

Stable Audio 3 Small SFX is a 459 million parameter latent diffusion model that generates high-quality sound effects from text prompts, designed for on-device deployment on mobile phones and consumer laptops.

sfxsound-effectson-device