EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 42 · 28 per page
text-to-image
falREVIEW REQUIRED

Bria Text-to-Image Base

fal-ai/bria/text-to-image/base

Bria's Text-to-Image model, trained exclusively on licensed data for safe and risk-free commercial use. Available also as source code and weights. For access to weights: https://bria.ai/contact-us

image generation
image-to-image
falREVIEW REQUIRED

Object Removal

fal-ai/image-apps-v2/object-removal

Remove unwanted objects seamlessly from any image.

removeobject-removal
image-to-image
falREVIEW REQUIRED

Florence-2 Large

fal-ai/florence-2-large/dense-region-caption

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

multimodalvision
video-to-video
moonvalleyREVIEW REQUIRED

Marey Realism V1.5

moonvalley/marey/motion-transfer

Pull motion from a reference video and apply it to new subjects or scenes.

video-to-video
AlibabaREVIEW REQUIRED

V2.6

wan/v2.6/reference-to-video/flash

Wan 2.6 reference-to-video flash model.

reference-to-video
training
MiniMaxREVIEW REQUIRED

MiniMax H3 Reference to Video LoRA Trainer

minimax/h3/ref2va/trainer

Train a MiniMax H3 LoRA with reference conditioning, so different modalities animate into video with audio; captions optional.

text-to-image
AlibabaREVIEW REQUIRED

Wan

fal-ai/wan/v2.2-5b/text-to-image

Wan 2.2's 5B model generates high-resolution, photorealistic images with powerful prompt understanding and fine-grained visual detail

text-to-image
falREVIEW REQUIRED

Bagel

fal-ai/bagel

Bagel is a 7B parameter from Bytedance-Seed multimodal model that can generate both text and images.

text-to-imagemultimodal
text-to-image
falREVIEW REQUIRED

Emu 3.5 Image

fal-ai/emu-3.5-image/text-to-image

Generate images from text using Emu 3.5 Image

image-to-3d
hitem3dREVIEW REQUIRED

Hi3D Multiview to 3D

hitem3d/hi3d/multi-view-to-3d

Generate 3D models from multiple view images using Hi3D.

image-to-3dmultiview-to-3d3d
audio-to-audio
falREVIEW REQUIRED

Stable Audio 3 Medium Audio Outpainting

fal-ai/stable-audio-3/medium/audio-outpainting

Stable Audio 3 Medium audio outpainting is a 1.4 billion parameter latent diffusion model that extends existing stereo audio beyond its original endpoint via causal continuation guided by text prompts.

musicextensioncontinuation
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit Plus Lora Gallery

fal-ai/qwen-image-edit-plus-lora-gallery/shirt-design

Apply designs/graphics onto people's shirts

stylizedtransform
image-to-video
MiniMaxREVIEW REQUIRED

MiniMax (Hailuo AI) Video 01 Director - Image to Video

fal-ai/minimax/video-01-director/image-to-video

Generate video clips more accurately with respect to initial image, natural language descriptions, and using camera movement instructions for shot control.

motiontransformationcamera-controls
text-to-video
falREVIEW REQUIRED

LongCat Video

fal-ai/longcat-video/text-to-video/480p

Generate long videos from text using LongCat Video

text-to-video
PixVerseREVIEW REQUIRED

PixVerse V4 Text To Video

fal-ai/pixverse/v4/text-to-video

Generate high quality video clips from text and image prompts using PixVerse v4

text-to-video
falREVIEW REQUIRED

Wan-2.1 Pro Text-to-Video

fal-ai/wan-pro/text-to-video

Wan-2.1 Pro is a premium text-to-video model that generates high-quality 1080p videos at 30fps with up to 6 seconds duration, delivering exceptional visual quality and motion diversity from text prompts

text to videomotion
image-to-image
falREVIEW REQUIRED

ControlNet SDXL

fal-ai/fast-sdxl-controlnet-canny/inpainting

Generate Images with ControlNet.

diffusioncontrolneteditingmanipulation
image-to-image
falREVIEW REQUIRED

Product Holding

fal-ai/image-apps-v2/product-holding

Place products naturally in a person’s hands for realistic marketing visuals.

productmarketing
image-to-3d
falREVIEW REQUIRED

Hunyuan3D

fal-ai/hunyuan3d/v2/mini/turbo

Generate 3D models from your images using Hunyuan 3D. A native 3D generative model enabling versatile and high-quality 3D asset creation.

stylized
image-to-image
falREVIEW REQUIRED

Finegrain Eraser Bbox

fal-ai/finegrain-eraser/bbox

Finegrain Eraser removes any object selected with a bounding box—along with its shadows, reflections, and lighting artifacts—seamlessly reconstructing the scene with contextually accurate content.

utilityediting
text-to-image
Black Forest LabsREVIEW REQUIRED

Juggernaut Flux Base

rundiffusion-fal/juggernaut-flux/base

Juggernaut Base Flux by RunDiffusion is a drop-in replacement for Flux [Dev] that delivers sharper details, richer colors, and enhanced realism, while instantly boosting LoRAs and LyCORIS with full compatibility.

image generation
image-to-image
falREVIEW REQUIRED

City Teleport

fal-ai/image-apps-v2/city-teleport

Place a person’s photo into iconic cities worldwide.

city-teleportbackgroundswap
image-to-image
falREVIEW REQUIRED

Texture Transform

fal-ai/image-apps-v2/texture-transform

Transform objects with different surface textures like marble, wood, or fabric.

texture-transform
video-to-video
falREVIEW REQUIRED

Lightx

fal-ai/lightx/recamera

Use the capabilities of lightx to relight and recamera your videos.

video-to-videorecamerarelight
image-to-image
briaREVIEW REQUIRED

Fibo Edit [Add Object by Text]

bria/fibo-edit/add_object_by_text

Precisely insert new objects into images with structured spatial commands. Context-aware, high-quality editing with seamless blending. Trained on licensed data for risk-free commercial and brand-safe use.

briafibo-editobject-additionjson
video-to-video
falREVIEW REQUIRED

ThinkSound

fal-ai/thinksound/audio

Generate realistic audio from a video with an optional text prompt

audio-generationvideo-to-audio
text-to-image
falREVIEW REQUIRED

Z Image Base Lora

fal-ai/z-image/base/lora

LoRA endpoint for Z-Image, the foundation model of the Z- Image family.

z-imagebaselora
image-to-video
PixVerseREVIEW REQUIRED

PixVerse V3.5 Transition

fal-ai/pixverse/v3.5/transition

Create seamless transition between images using PixVerse v3.5