EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 17 · 28 per page
image-to-3d
hitem3dREVIEW REQUIRED

Hi3D Multiview to 3D

hitem3d/hi3d/v3.0/multi-view-to-3d

Generate 3D models from multiple view images using Hi3D V3.0.

image-to-3dmultiview-to-3d3d
vision
falREVIEW REQUIRED

Florence-2 Large

fal-ai/florence-2-large/more-detailed-caption

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

captioningmultimodalvision
image-to-image
falREVIEW REQUIRED

Florence-2 Large

fal-ai/florence-2-large/object-detection

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

detectionmultimodalvision
text-to-image
falREVIEW REQUIRED

Stable Diffusion 3.5 Large

fal-ai/stable-diffusion-v35-large

Stable Diffusion 3.5 Large is a Multimodal Diffusion Transformer (MMDiT) text-to-image model that features improved performance in image quality, typography, complex prompt understanding, and resource-efficiency.

diffusiontypographystyle
image-to-image
falREVIEW REQUIRED

Feynobg Background Remover

fal-ai/feynobg

FeyNobg is a state of the art AI model for background removal from feyninc

utilityediting
image-to-image
falREVIEW REQUIRED

Smart Resize

fal-ai/smart-resize

Smart image resize to arbitrary dimensions, powered by Nano Banana Pro with vision-LLM-guided prompting for composition-aware recomposition. Crop, cropping, resize ads.

realismtypographyvisualads
image-to-video
falREVIEW REQUIRED

Pika Image to Video (v2.2)

fal-ai/pika/v2.2/image-to-video

Turn photos into mind-blowing, dynamic videos in up to 1080p. Experience better image clarity and crisper, sharper visuals.

editingeffectsanimation
image-to-image
falREVIEW REQUIRED

Finegrain Eraser Mask

fal-ai/finegrain-eraser/mask

Finegrain Eraser removes any object selected with a mask—along with its shadows, reflections, and lighting artifacts—seamlessly reconstructing the scene with contextually accurate content.

utilityediting
image-to-video
AlibabaREVIEW REQUIRED

Happy Horse 1.1 Reference to Video

alibaba/happy-horse/v1.1/reference-to-video

Happy Horse 1.1 is Alibaba's #1-ranked video model. This reference-to-video endpoint turns up to 9 reference images into 1080p video with synchronized native audio and multilingual lip-sync for consistent characters.

happy-horsevideoreference
text-to-speech
ByteDanceREVIEW REQUIRED

Bytedance Seed Speech Text to Speech

fal-ai/bytedance/seed-speech/tts/v2

Seed Speech developed by ByteDance, is a family of large-scale text-to-speech models capable of synthesizing speech that is virtually indistinguishable from human speech.

stylizedtransformlipsync
text-to-3d
MeshyREVIEW REQUIRED

V7 Text to 3D

meshy/v7/text-to-3d

Turns text into a fully textured, PBR-ready 3D mesh with complete geometry, in game-ready Smart Topology at a target polygon count

stylizedtransform
image-to-image
IdeogramREVIEW REQUIRED

Ideogram V3 Edit

fal-ai/ideogram/v3/edit

Transform existing images with Ideogram V3's editing capabilities. Modify, adjust, and refine images while maintaining high fidelity and realistic outputs with precise prompt control.

realismtypography
image-to-image
RecraftREVIEW REQUIRED

Recraft V3

fal-ai/recraft/v3/image-to-image

Recraft V3 is a text-to-image model with the ability to generate long texts, vector art, images in brand style, and much more. As of today, it is SOTA in image generation, proven by Hugging Face's industry-leading Text-to-Image Benchmark by Artificial Analysis.

vectortypographystyle
image-to-video
PixVerseREVIEW REQUIRED

PixVerse V6 Transition

fal-ai/pixverse/v6/transition

Pixverse's latest v6 Model.

image-to-videofirst-frame-last-frametransition
image-to-video
PixVerseREVIEW REQUIRED

PixVerse C1 Image To Video

fal-ai/pixverse/c1/image-to-video

Animate images into cinematic videos with PixVerse C1, supporting 1080p resolution and native audio generation.

video-generationimage-to-videopixverseanimation
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev] with Controlnets and Loras

fal-ai/flux-general/image-to-image

FLUX General Image-to-Image is a versatile endpoint that transforms existing images with support for LoRA, ControlNet, and IP-Adapter extensions, enabling precise control over style transfer, modifications, and artistic variations through multiple guidance methods.

loracontrolnetip-adapter
text-to-3d
tripo3dREVIEW REQUIRED

Tripo H3.1 Text to 3D

tripo3d/h3.1/text-to-3d

Generate 3D models from text descriptions using Tripo H3.1.

3dtext-to-3d3d-generationtripo
image-to-image
OpenAIREVIEW REQUIRED

GPT Image 1 Mini

fal-ai/gpt-image-1-mini/edit

GPT Image 1 mini combines OpenAI's advanced language capabilities, powered by GPT-5, with GPT Image 1 Mini for efficient image generation.

image-to-image
vision
falREVIEW REQUIRED

Moondream2

fal-ai/moondream2/visual-query

Moondream2 is a highly efficient open-source vision language model that combines powerful image understanding capabilities with a remarkably small footprint.

Vision
image-to-video
Luma AIREVIEW REQUIRED

Luma Ray 2 (Image to Video)

fal-ai/luma-dream-machine/ray-2/image-to-video

Ray2 is a large-scale video generative model capable of creating realistic visuals with natural, coherent motion.

motiontransformation
text-to-speech
resemble-aiREVIEW REQUIRED

Chatterboxhd

resemble-ai/chatterboxhd/text-to-speech

Generate expressive, natural speech with Resemble AI's Chatterbox. Features unique emotion control, instant voice cloning from short audio, and built-in watermarking.

text-to-audio
MiniMaxREVIEW REQUIRED

MiniMax (Hailuo AI) Music

fal-ai/minimax-music

Generate music from text prompts using the MiniMax model, which leverages advanced AI techniques to create high-quality, diverse musical compositions.

music
image-to-image
falREVIEW REQUIRED

Virtual Try-on

fal-ai/image-apps-v2/virtual-try-on

Try on clothes virtually by combining person and clothing images.

fashiontry-onvirtual-try-on
text-to-video
falREVIEW REQUIRED

LTX Video 2.3 Pro

fal-ai/ltx-2.3/text-to-video

LTX-2.3 is a high-quality, fast AI video model available in Pro and Fast variants for text-to-video, image-to-video, and audio-to-video.

stylizedtransformlipsync
video-to-video
Black Forest LabsREVIEW REQUIRED

Flux 3 Extend Video

blackforestlabs/flux-3/extend-video

FLUX 3 is Black Forest Labs' frontier video model. This endpoint continues an existing clip beyond its final frame, generating additional footage that stays consistent with the original motion and scene.

stylizedtransformlipsync
text-to-image
falREVIEW REQUIRED

Z Image Base

fal-ai/z-image/base

Z-Image is the foundation model of the Z- Image family, engineered for good quality, robust generative diversity, broad stylistic coverage, and precise prompt adherence.

z-imagebase