EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 13 · 28 per page
image-to-video
AlibabaREVIEW REQUIRED

Happy Horse 1.1 Image to Video

alibaba/happy-horse/v1.1/image-to-video

Happy Horse 1.1 is Alibaba's #1-ranked video model. This image-to-video endpoint animates a still image into 1080p video with synchronized native audio and multilingual lip-sync

happy-horsevideoimage
text-to-video
xAIREVIEW REQUIRED

Grok Imagine Video 1.5 Text to Video

xai/grok-imagine-video/v1.5/text-to-video

Generate videos from prompts with audio using xAI's Grok Imagine 1.5 Video model.

stylizedtransformlipsync
text-to-video
MiniMaxREVIEW REQUIRED

MiniMax Hailuo 02 [Standard] (Text to Video)

fal-ai/minimax/hailuo-02/standard/text-to-video

MiniMax Hailuo-02 Text To Video API (Standard, 768p): Advanced video generation model with 768p resolution

text-to-image
RecraftREVIEW REQUIRED

Recraft V4 Pro

fal-ai/recraft/v4/pro/text-to-image

Recraft V4 was developed with designers to bring true visual taste to AI image generation. Built for brand systems and production-ready workflows, it goes beyond prompt accuracy — delivering stronger composition, refined lighting, realistic materials, and a cohesive aesthetic. The result is imagery shaped by professional design judgment, ready for immediate real-world use without additional post-processing.

text-to-image
image-to-image
falREVIEW REQUIRED

Image Outpaint

fal-ai/image-apps-v2/outpaint

Directional outpainting. Choose edges to expand. left, right, top, or center (uniform all sides). Only expanded areas are generated; an optional zoom-out pulls the frame back by the chosen amount.

outpainting
image-to-json
briaREVIEW REQUIRED

Bria Ad Delayer: Convert Flat Ads into Editable Layers | fal

bria/ad-delayer

Turn any flat ad image into fully editable layers —background, product and logo cutouts, live text with typography, and vector shapes. Commercial-safe, structured JSON output

utilityediting
text-to-video
falREVIEW REQUIRED

Wan 2.5 Text to Video

fal-ai/wan-25-preview/text-to-video

Wan 2.5 text-to-video model.

image-to-3d
falREVIEW REQUIRED

Sam 3

fal-ai/sam-3/3d-objects

SAM 3D enables precise 3D reconstruction of objects from real images, while accurately reconstructing their geometry and texture.

3dobject
3d-to-3d
MeshyREVIEW REQUIRED

Meshy Rigging

fal-ai/meshy/rigging

Rig humanoid 3D models from GLB URLs with Meshy, returning rigged GLB/FBX files plus basic animations.

3d-to-3drigging
image-to-image
microsoftREVIEW REQUIRED

Mai Image 2.5

microsoft/mai-image-2.5/edit

MAI-Image-2.5 is Microsoft's photorealistic image generation and editing model that turns text prompts or uploaded images into high-quality, design-ready visuals with fine-grained, pixel-level control.

realismtypographystylized
image-to-video
falREVIEW REQUIRED

Creatify Aurora

fal-ai/creatify/aurora

Generate high fidelity, studio quality videos of your avatar speaking or singing using the Aurora from Creatify team!

lipsyncimage-to-video
text-to-image
kreaREVIEW REQUIRED

Krea 2 Medium

krea/v2/medium/text-to-image

Generate high-quality images from text with Krea 2 Medium, supporting aspect ratio, creativity controls, seeds, and optional style references.

text-to-imageimage-generationstyle-referencekrea
video-to-video
AlibabaREVIEW REQUIRED

Happy Horse Video Edit

alibaba/happy-horse/video-edit

HappyHorse video editing supports advanced video editing through natural language instructions. It allows for local or global editing of video elements using up to 5 reference images.

happy-horsevideo-editingvideo-to-video
image-to-video
AlibabaREVIEW REQUIRED

Wan v2.2 5B

fal-ai/wan/v2.2-5b/image-to-video

Wan 2.2's 5B model produces up to 5 seconds of video 720p at 24FPS with fluid motion and powerful prompt understanding

text-to-video
falREVIEW REQUIRED

LTX 2.3 Video Fast

fal-ai/ltx-2.3/text-to-video/fast

LTX-2.3 is a high-quality, fast AI video model available in Pro and Fast variants for text-to-video, image-to-video, and audio-to-video.

stylizedtransformlipsync
text-to-image
falREVIEW REQUIRED

PATINA

fal-ai/patina/material

Generate complete seamlessly tiling PBR materials including normal, roughness, basecolor, height and metalness maps up to 8K

materialpbrdisplacementmetalness
image-to-image
IdeogramREVIEW REQUIRED

Ideogram V3 Character

fal-ai/ideogram/character

Generate consistent character appearances across multiple images. Maintain facial features, proportions, and distinctive traits for cohesive storytelling and branding

character-consistency
image-to-image
topazREVIEW REQUIRED

Topaz Upscale Image Creative

topaz/upscale/image/creative

Professional creative image upscaling powered by Topaz Labs. Bloom 2 reinvents detail with adjustable creativity and color preservation. Best for AI-generated images that need striking enhancement.

upscaleimage
audio-to-audio
falREVIEW REQUIRED

ACE Step Audio To Audio

fal-ai/ace-step/audio-to-audio

Generate music from a lyrics and example audio using ACE-Step

audio-to-audioaudio-edit
text-to-video
Black Forest LabsREVIEW REQUIRED

Flux 3 Text To Video Draft

blackforestlabs/flux-3/text-to-video/draft

FLUX.3 is Black Forest Labs' frontier audio/video model. Generate fast, low-cost draft previews from a text prompt, with a reusable draft cache for full-quality enhancement.

stylizedtransformlipsync
image-to-3d
MeshyREVIEW REQUIRED

Meshy 6

fal-ai/meshy/v6/image-to-3d

Meshy-6 is the latest model from Meshy. It generates realistic and production ready 3D models.

image-to-3d
text-to-image
KlingREVIEW REQUIRED

Kling Image

fal-ai/kling-image/o3/text-to-image

Kling Omni 3: Top-tier text-to-image with flawless consistency.

text-to-image
audio-to-audio
AlibabaREVIEW REQUIRED

Qwen 3 TTS - Clone Voice [1.7B]

fal-ai/qwen-3-tts/clone-voice/1.7b

Clone your voices using Qwen3-TTS Clone-Voice model with zero shot cloning capabilities and use it on text-to-speech models to create speeches of yours!

clone-voicevoice-clone
image-to-image
falREVIEW REQUIRED

Hunyuan Image

fal-ai/hunyuan-image/v3/instruct/edit

Image editing endpoint for Hunyuan Image 3.0 Instruct.

tencenthunyuan-imageinstructedit
image-to-3d
falREVIEW REQUIRED

Hyper3D Rodin

fal-ai/hyper3d/rodin

Rodin by Hyper3D generates realistic and production ready 3D models from text or images.

stylized
text-to-video
KlingREVIEW REQUIRED

Kling Video V3 Turbo Pro Text to Video

fal-ai/kling-video/v3/turbo/pro/text-to-video

Generate high quality 1080p videos using Kling's Turbo 3.0 model, with improved lipsync and multishot generation capabilities.

klingv31080pturbo
text-to-image
MiniMaxREVIEW REQUIRED

MiniMax (Hailuo AI) Text to Image

fal-ai/minimax/image-01

Generate high quality images from text prompts using MiniMax Image-01. Longer text prompts will result in better quality images.

stylizedrealism
text-to-image
falREVIEW REQUIRED

Stable Diffusion XL Lightning

fal-ai/fast-lightning-sdxl

Run SDXL at the speed of light

diffusionlightningreal-time