EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 24 · 28 per page
text-to-3d
falREVIEW REQUIRED

Hyper3D - Rodin V2.5 - Text to 3D

fal-ai/hyper3d/rodin/v2.5/text-to-3d

Rodin V2.5 by Hyper3D generates realistic and production ready 3D models from text or images.

text-to-3d
image-to-image
topazREVIEW REQUIRED

Topaz Sharpen Image

topaz/sharpen/image

Professional photo sharpening powered by Topaz Labs. Models tuned per blur type (lens, motion, portrait, wildlife), plus Super Focus for generative recovery of severely blurred shots. Best for out-of-focus and motion-blurred photos.

sharpenimage
image-to-video
AlibabaREVIEW REQUIRED

Wan v2.2 A14B Image-to-Video A14B with LoRAs

fal-ai/wan/v2.2-a14b/image-to-video/lora

Wan-2.2 image-to-video is a video model that generates high-quality videos with high visual quality and motion diversity from text prompts and images. This endpoint supports LoRAs made for Wan 2.2

image-to-videomotionlora
audio-to-video
AlibabaREVIEW REQUIRED

Wan-2.2 Speech-to-Video 14B

fal-ai/wan/v2.2-14b/speech-to-video

Wan-S2V is a video model that generates high-quality videos from static images and audio, with realistic facial expressions, body movements, and professional camera work for film and television applications

audio-to-videotalking-head
image-to-image
briaREVIEW REQUIRED

Extract Object

bria/extract-object

Bria Extract Object uses text prompts to isolate a selected object from an image and return it as an RGBA PNG with a transparent background. Ideal for product, ecommerce, advertising, and creative editing workflows. Bria's Extract Object API leads in product shot extraction, outperforming SAM 3.1 where it counts most for commercial use.

text-to-image
falREVIEW REQUIRED

Wan 2.5 Text to Image

fal-ai/wan-25-preview/text-to-image

Wan 2.5 text-to-image model.

video-to-video
Luma AIREVIEW REQUIRED

Luma Ray 2 Modify

fal-ai/luma-dream-machine/ray-2/modify

Ray2 Modify is a video generative model capable of restyling or retexturing the entire shot, from turning live-action into CG or stylized animation, to changing wardrobe, props, or the overall aesthetic and swap environments or time periods, giving you control over background, location, or even weather.

modifyrestyle
video-to-video
KlingREVIEW REQUIRED

Kling O1 Edit Video [Standard]

fal-ai/kling-video/o1/standard/video-to-video/edit

Edit an existing video using natural-language instructions, transforming subjects, settings, and style while retaining the original motion structure.

text-to-audio
falREVIEW REQUIRED

Kokoro TTS (British English)

fal-ai/kokoro/british-english

A high-quality British English text-to-speech model offering natural and expressive voice synthesis.

speech
video-to-video
mirelo-aiREVIEW REQUIRED

Mirelo SFX1.6

mirelo-ai/sfx1.6/video-to-video

Generate synced sounds for any video, and return it with its new sound track (like MMAudio). Now up to 60 seconds!

video-to-videosfx
text-to-image
imagineartREVIEW REQUIRED

ImagineArt 1.5 Pro Preview

imagineart/imagineart-1.5-pro-preview/text-to-image

ImagineArt 1.5 Pro is an advanced text-to-image model that creates ultra-high-fidelity 4K visuals with lifelike realism, refined aesthetics, and powerful creative output suited for professional use.

visualsimagineartrealismtext
text-to-image
falREVIEW REQUIRED

Stable Diffusion V3

fal-ai/stable-diffusion-v3-medium

Stable Diffusion 3 Medium (Text to Image) is a Multimodal Diffusion Transformer (MMDiT) model that improves image quality, typography, prompt understanding, and efficiency.

diffusionstyle
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev] Control LoRA Canny

fal-ai/flux-control-lora-canny

FLUX Control LoRA Canny is a high-performance endpoint that uses a control image to transfer structure to the generated image, using a Canny edge map.

lorastyle transfer
audio-to-audio
falREVIEW REQUIRED

Stable Audio 3 Medium Audio to Audio

fal-ai/stable-audio-3/medium/audio-to-audio

Stable Audio 3 Medium audio-to-audio is a 1.4 billion parameter latent diffusion model that transforms an input audio clip into new stereo variations up to 6 minutes guided by a text prompt.

musicstyle-transferremix
text-to-video
falREVIEW REQUIRED

LTX Video-0.9.7 13B Distilled

fal-ai/ltx-video-13b-distilled

Generate videos from prompts using LTX Video-0.9.7 13B Distilled and custom LoRA

videoltx-videotext-to-video
text-to-image
Luma AIREVIEW REQUIRED

Luma Uni-1 Text to Image Max

luma/agent/uni-1/v1/max

Luma Uni-1 Max generates a single image at the model's highest fidelity, delivering richer detail and stronger prompt adherence than the base tier for hero-quality stills.

realismtypographystylized
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.2 [klein] 9B Base LoRA

fal-ai/flux-2/klein/9b/base/lora

Text-to-image generation with LoRA support for FLUX.2 [klein] 9B Base from Black Forest Labs. Custom style adaptation and fine-tuned model variations.

text-to-image
Black Forest LabsREVIEW REQUIRED

Juggernaut Flux Lightning

rundiffusion-fal/juggernaut-flux/lightning

Juggernaut Lightning Flux by RunDiffusion provides blazing-fast, high-quality images rendered at five times the speed of Flux. Perfect for mood boards and mass ideation, this model excels in both realism and prompt adherence.

image generation
audio-to-video
falREVIEW REQUIRED

LTX 2.3 Video Pro

fal-ai/ltx-2.3/audio-to-video

LTX-2.3 is a high-quality, fast AI video model available in Pro and Fast variants for text-to-video, image-to-video, and audio-to-video.

stylizedtransformlipsync
text-to-video
Luma AIREVIEW REQUIRED

Luma Ray 2

fal-ai/luma-dream-machine/ray-2

Ray2 is a large-scale video generative model capable of creating realistic visuals with natural, coherent motion.

motiontransformation
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX1.1 [pro] ultra Redux

fal-ai/flux-pro/v1.1-ultra/redux

FLUX1.1 [pro] ultra Redux is a high-performance endpoint for the FLUX1.1 [pro] model that enables rapid transformation of existing images, delivering high-quality style transfers and image modifications with the core FLUX capabilities.

style transferhigh-res
image-to-image
falREVIEW REQUIRED

Image Editing Object Removal

fal-ai/image-editing/object-removal

Remove unwanted objects or people from your photos while seamlessly blending the background.

stylizedtransform
video-to-video
falREVIEW REQUIRED

Sync React-1

fal-ai/sync-lipsync/react-1

Use React-1 from SyncLabs to refine human emotions and do realistic lip-sync without losing details!

lipsyncvideo-to-video
llm
ByteDanceREVIEW REQUIRED

Bytedance Seed V2 Mini

fal-ai/bytedance/seed/v2/mini

Seed 2.0 Mini is a high-performance multimodal model optimized for low latency and high concurrency. It supports text, image, and video input with 256K context and configurable thinking/reasoning modes.

text-to-video
KlingREVIEW REQUIRED

Kling Video

fal-ai/kling-video/o3/4k/text-to-video

Kling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling

stylizedtransformlipsync
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.2 [klein] 4B LoRA

fal-ai/flux-2/klein/4b/edit/lora

Image-to-image editing with FLUX.2 [klein] 4B from Black Forest Labs and custom LoRA. Precise modifications using natural language descriptions and hex color control.