EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 22 · 28 per page
llm
openrouterREVIEW REQUIRED

OpenRouter [Enterprise]

openrouter/router/enterprise

Run any LLM (Large Language Model) with fal, powered by OpenRouter.

vision
falREVIEW REQUIRED

MoonDreamNext

fal-ai/moondream-next

MoonDreamNext is a multimodal vision-language model for captioning, gaze detection, bbox detection, point detection, and more.

multimodalvision
image-to-image
falREVIEW REQUIRED

Image2svg

fal-ai/image2svg

Image2SVG transforms raster images into clean vector graphics, preserving visual quality while enabling scalable, customizable SVG outputs with precise control over detail levels.

utilityediting
text-to-image
RecraftREVIEW REQUIRED

Recraft V4 Styles Pro Text to Image

recraft/v4/style/pro/text-to-image

Generates raster images that hold a consistent style, from either a saved style ID or reference images attached directly.

stylizedtransformediting
text-to-image
RecraftREVIEW REQUIRED

Recraft V4.1 Text to Vector Pro

fal-ai/recraft/v4.1/pro/text-to-vector

Recraft V4.1 Pro Vector generates large-format, fully editable SVGs with the structural clarity professional illustrators expect. Built for poster art, complex brand assets, and detailed scene illustration, it scales without losing geometric integrity.

stylizedtransformtypography
image-to-image
falREVIEW REQUIRED

Hidream O1 Image

fal-ai/hidream-o1-image/edit

Unified image generation with HiDream-O1-Image. Create, edit, and personalize high-resolution images up to 2K—single native model handles text-to-image, editing, and custom subjects without external components.

image-to-video
Luma AIREVIEW REQUIRED

Luma Ray 2 Flash (Image to Video)

fal-ai/luma-dream-machine/ray-2-flash/image-to-video

Ray2 Flash is a fast video generative model capable of creating realistic visuals with natural, coherent motion.

motiontransformation
text-to-video
MiniMaxREVIEW REQUIRED

MiniMax Hailuo 2.3 [Pro] (Text to Video)

fal-ai/minimax/hailuo-2.3/pro/text-to-video

MiniMax Hailuo-2.3 Text To Video API (Pro, 1080p): Advanced text-to-video generation model with 1080p resolution

text-to-video
image-to-image
Black Forest LabsREVIEW REQUIRED

Flux 2 [klein] Realtime

fal-ai/flux-2/klein/realtime

Realtime generation with FLUX.2 [klein] from Black Forest Labs.

realtimeimage-to-image
text-to-audio
mirelo-aiREVIEW REQUIRED

Mirelo SFX1.6

mirelo-ai/sfx1.6/text-to-audio

Generate ambient sounds for any text prompt. Now you can turn any SFX into a natural loop for ambient soundscapes.

text-to-audiosfx
image-to-image
Luma AIREVIEW REQUIRED

Luma Uni-1 Edit Max

luma/agent/uni-1/v1/max/edit

Luma Uni-1 Max Edit applies text-guided edits to a source image at maximum fidelity, holding the original structure while honoring reference images for precise, high-detail revisions.

realismtypographystylized
text-to-audio
MiniMaxREVIEW REQUIRED

Minimax Music 2.5

fal-ai/minimax-music/v2.5

MiniMax Music 2.5 creates complete tracks with singing, backing music, and detailed arrangements from lyrics and a style description.

stylizedtransformlipsync
video-to-video
KlingREVIEW REQUIRED

Kling O1 Reference Video to Video [Pro]

fal-ai/kling-video/o1/video-to-video/reference

Kling O1 Omni generates new shots guided by an input reference video, preserving cinematic language such as motion, and camera style to produce seamless scene continuity.

video-to-video
falREVIEW REQUIRED

Wan Motion

fal-ai/wan-motion

Wan Motion is a streamlined character animation model that transfers motion from a driving video onto a reference character image. Based on Wan-Animate which preserves the original character's proportions, Simple uses pose retargeting to adapt the driving video's skeleton to match the reference character's body shape, producing more natural results when the two have different builds. It outputs at 720p with optimized defaults for fast, high-quality generation — just provide a video, an image, and an optional prompt.

text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 Krea [dev]

fal-ai/flux-1/krea

FLUX.1 Krea [dev] is a 12 billion parameter flow transformer that generates high-quality images from text with incredible aesthetics. It is suitable for personal and commercial use.

vision
falREVIEW REQUIRED

Moondream2

fal-ai/moondream2

Moondream2 is a highly efficient open-source vision language model that combines powerful image understanding capabilities with a remarkably small footprint.

Vision
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.2 [klein] 9B Base

fal-ai/flux-2/klein/9b/base

Text-to-image generation with FLUX.2 [klein] 9B Base from Black Forest Labs. Enhanced realism, crisper text generation, and native editing capabilities.

vision
falREVIEW REQUIRED

Moondream2

fal-ai/moondream2/object-detection

Moondream2 is a highly efficient open-source vision language model that combines powerful image understanding capabilities with a remarkably small footprint.

image-to-image
text-to-image
falREVIEW REQUIRED

Ernie Image Turbo

fal-ai/ernie-image/turbo

High-quality text-to-image model by Baidu. Supports English, Chinese, and Japanese prompts with built-in prompt expansion.

image-to-video
PixVerseREVIEW REQUIRED

PixVerse C1 Reference To Video

fal-ai/pixverse/c1/reference-to-video

Generate character-consistent videos from reference images using PixVerse C1, with subject and background references.

video-generationreference-to-videopixversecharacter-consistency
video-to-video
GoogleREVIEW REQUIRED

Veo 3.1

fal-ai/veo3.1/extend-video

Extend Veo-Created Videos up to 30 seconds

extend-video
image-to-image
falREVIEW REQUIRED

FASHN Virtual Try-On V1.5

fal-ai/fashn/tryon/v1.5

FASHN v1.5 delivers precise virtual try-on capabilities, accurately rendering garment details like text and patterns at 576x864 resolution from both on-model and flat-lay photo references.

try-onfashionclothing
text-to-image
falREVIEW REQUIRED

Hidream I1 Full

fal-ai/hidream-i1-full

HiDream-I1 full is a new open-source image generative foundation model with 17B parameters that achieves state-of-the-art image generation quality within seconds.

image-to-image
falREVIEW REQUIRED

Marigold Depth Estimation

fal-ai/imageutils/marigold-depth

Create depth maps using Marigold depth estimation.

depthutility
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX Virtual Try-On

fal-ai/flux-pro/v1/vto

Generate virtual try-on results from a person image plus one or more garment references.

image-to-imagevton
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX 2 Lora Gallery Realism

fal-ai/flux-2-lora-gallery/realism

Makes images more photorealistic and natural

stylizedtransform
text-to-image
briaREVIEW REQUIRED

Fibo Gen 1.5 Text to Image

bria/fibo-gen-1.5/text-to-image

Text-to-image model with high-fidelity outputs, accurate typography, and style preset, strong in photorealism, textures, and beyond. JSON-structured prompts give enterprise and agentic workflows production-ready control. Trained on licensed data.

stylizedtransformrealism
image-to-image
falREVIEW REQUIRED

Z Image Turbo Image To Image Lora

fal-ai/z-image/turbo/image-to-image/lora

Generate images from text and images using custom LoRA and Z-Image Turbo, Tongyi-MAI's super-fast 6B model.

turboz-imagefastlora