EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 34 · 28 per page
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 Subject

fal-ai/flux-subject

Super fast endpoint for the FLUX.1 [schnell] model with subject input capabilities, enabling rapid and high-quality image generation for personalization, specific styles, brand identities, and product-specific outputs.

personalizationcustomization
video-to-video
falREVIEW REQUIRED

Wan VACE 14B

fal-ai/wan-vace-14b/depth

VACE is a video generation model that uses a source image, mask, and video to create prompted videos with controllable sources.

image-to-videovideo-to-videotext-to-video
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 Krea [dev] Inpainting with LoRAs

fal-ai/flux-krea-lora/inpainting

Super fast endpoint for the FLUX.1 [dev] inpainting model with LoRA support, enabling rapid and high-quality image inpaingting using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.

lorapersonalization
text-to-speech
falREVIEW REQUIRED

VibeVoice 1.5B

fal-ai/vibevoice

Generate long, expressive multi-voice speech using Microsoft's powerful TTS

text-to-speechmulti-speakerpodcast
text-to-speech
asyncREVIEW REQUIRED

Async Text to Speech Pro V1.0

async/tts-pro/v1.0

Generate professional-quality voiceovers in seconds with Async TTS Pro model text-based control over pauses, emphasis, and timing. Voice ids can be found at https://async.com/developer/voice-library

text-to-speechvoice-clonelipsync
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit Plus Lora Gallery

fal-ai/qwen-image-edit-plus-lora-gallery/remove-element

Remove unwanted elements (objects, people, text) while maintaining image consistency

stylizedtransform
image-to-3d
falREVIEW REQUIRED

Hunyuan World

fal-ai/hunyuan_world/image-to-world

Hunyuan World 1.0 turns a single image into a panorama or a 3D world. It creates realistic scenes from the image, allowing you to explore and view it from different angles.

image-to-video
nvidiaREVIEW REQUIRED

Cosmos 3 Super Image to Video

nvidia/cosmos-3-super/image-to-video

Cosmos3 is a collection of Omnimodal world models capable of generating dynamic, high-quality video, image, audio, and action commands from combinations of text, image, video, and action trajectory inputs.

stylizedtransformlipsync
text-to-video
falREVIEW REQUIRED

LTX-2 19B Distilled

fal-ai/ltx-2-19b/distilled/text-to-video

Generate video with audio from text using LTX-2 Distilled

video-to-video
falREVIEW REQUIRED

Wan 2.1 VACE Long Reframe

fal-ai/wan-vace-apps/long-reframe

Reframe entire videos scene-by-scene using Wan VACE 2.1

image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX1.1 [pro] Redux

fal-ai/flux-pro/v1.1/redux

FLUX1.1 [pro] Redux is a high-performance endpoint for the FLUX1.1 [pro] model that enables rapid transformation of existing images, delivering high-quality style transfers and image modifications with the core FLUX capabilities.

style transfer
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.2 [klein] 4B Base LoRA

fal-ai/flux-2/klein/4b/base/lora

Text-to-image generation with LoRA support for FLUX.2 [klein] 4B Base from Black Forest Labs. Custom style adaptation and fine-tuned model variations.

text-to-image
Black Forest LabsREVIEW REQUIRED

Flux Kontext Lora

fal-ai/flux-kontext-lora/text-to-image

Super fast text-to-image endpoint for the FLUX.1 Kontext [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.

text-to-image
video-to-audio
falREVIEW REQUIRED

Sam Audio

fal-ai/sam-audio/visual-separate

Audio separation with SAM Audio. Isolate any sound using natural language—professional-grade audio editing made simple for creators, researchers, and accessibility applications.

video-to-audiosam-audio
text-to-image
falREVIEW REQUIRED

Hidream I1 Dev

fal-ai/hidream-i1-dev

HiDream-I1 dev is a new open-source image generative foundation model with 17B parameters that achieves state-of-the-art image generation quality within seconds.

video-to-video
falREVIEW REQUIRED

Ltx 2.3 Quality

fal-ai/ltx-2.3-quality/extend-video

Extend high-quality video with audio from input video using LTX-2.3

extendlonger
text-to-image
Luma AIREVIEW REQUIRED

Luma Photon Flash

fal-ai/luma-photon/flash

Generate images from your prompts using Luma Photon Flash. Photon Flash is the most creative, personalizable, and intelligent visual models for creatives, bringing a step-function change in the cost of high-quality image generation.

video-to-video
falREVIEW REQUIRED

Infinitalk

fal-ai/infinitalk/video-to-video

Infinitalk model generates a talking avatar video from an image and audio file. The avatar lip-syncs to the provided audio with natural facial expressions.

video-to-video
text-to-image
falREVIEW REQUIRED

DeepSeek Janus-Pro

fal-ai/janus

DeepSeek Janus-Pro is a novel text-to-image model that unifies multimodal understanding and generation through an autoregressive framework

stylized
text-to-image
falREVIEW REQUIRED

PixArt-Σ

fal-ai/pixart-sigma

Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

diffusion
text-to-image
falREVIEW REQUIRED

Longcat Image

fal-ai/longcat-image

LongCat image is a 6B parameter model excelling at multilingual text rendering, photorealism and deployment efficiency.

image-to-image
falREVIEW REQUIRED

Bagel

fal-ai/bagel/edit

Bagel is a 7B parameter multimodal model from Bytedance-Seed that can generate both images and text.

image-to-imageimage-editing
image-to-image
falREVIEW REQUIRED

Perspective Change

fal-ai/image-apps-v2/perspective

Easily adjust the perspective of any image to different angles.

change-angleperspective
image-to-image
IdeogramREVIEW REQUIRED

Ideogram V4.0q Tiling

ideogram/v4/tiling

Ideogram V4.0q Tiling generates seamless, edge-matching textures and patterns that repeat infinitely in any direction, ideal for backgrounds, surfaces, and wallpapers.

stylizedtransformrealism
image-to-image
falREVIEW REQUIRED

Ghiblify Images

fal-ai/ghiblify

Reimagine and transform your ordinary photos into enchanting Studio Ghibli style artwork

stylizedtransform
vision
falREVIEW REQUIRED

GOT OCR 2.0

fal-ai/got-ocr/v2

GOT-OCR2 works on a wide range of tasks, including plain document OCR, scene text OCR, formatted document OCR, and even OCR for tables, charts, mathematical formulas, geometric shapes, molecular formulas and sheet music.

optical character recognitionhigh-resutility
text-to-image
Black Forest LabsREVIEW REQUIRED

Flux 2 Lora Gallery

fal-ai/flux-2-lora-gallery/digital-comic-art

Transforms images into comic book style

stylizedtransform