EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 12 · 28 per page
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev] with Controlnets and Loras

fal-ai/flux-general

A versatile endpoint for the FLUX.1 [dev] model that supports multiple AI extensions including LoRA, ControlNet conditioning, and IP-Adapter integration, enabling comprehensive control over image generation through various guidance methods.

loracontrolnetip-adapter
image-to-image
falREVIEW REQUIRED

SeedVR2

fal-ai/seedvr/upscale/image/seamless

Use SeedVR2 to upscale images, retaining seamless tiling

upscaleimage-to-imageseamlesstiling
text-to-speech
MiniMaxREVIEW REQUIRED

MiniMax Voice Design

fal-ai/minimax/voice-design

Design a personalized voice from a text description, and generate speech from text prompts using the MiniMax model, which leverages advanced AI techniques to create high-quality text-to-speech.

speech
image-to-image
falREVIEW REQUIRED

CodeFormer

fal-ai/codeformer

Fix distorted or blurred photos of people with CodeFormer.

image-restorationfacesutility
image-to-video
falREVIEW REQUIRED

LTX 2.3 Video Pro

fal-ai/ltx-2.3/image-to-video

LTX-2.3 is a high-quality, fast AI video model available in Pro and Fast variants for text-to-video, image-to-video, and audio-to-video.

stylizedtransformlipsync
image-to-video
GoogleREVIEW REQUIRED

Veo3.1 Lite FLF

fal-ai/veo3.1/lite/first-last-frame-to-video

Veo 3.1 Lite balances practical utility with professional capabilities, supporting Text-to-Video and Image-to-Video

stylizedtransformlipsync
image-to-image
AlibabaREVIEW REQUIRED

Wan

fal-ai/wan/v2.7/pro/edit

Edit and transform images using text instructions with the WAN 2.7 Pro model for precise, professional-grade image modifications.

wanimage-editingpro
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX 2 Lora Edit

fal-ai/flux-2/lora/edit

Image-to-image editing with LoRA support for FLUX.2 [dev] from Black Forest Labs. Specialized style transfer and domain-specific modifications.

text-to-audio
falREVIEW REQUIRED

ACE Step Prompt To Audio

fal-ai/ace-step/prompt-to-audio

Generate music from a simple prompt using ACE-Step

text-to-audiotext-to-music
text-to-image
AlibabaREVIEW REQUIRED

Qwen Image 2512

fal-ai/qwen-image-2512

Qwen Image 2512 is an improved version of Qwen Image with better text rendering, finer natural textures, and more realistic human generation.

qwen2512
image-to-3d
falREVIEW REQUIRED

Hunyuan3d V3

fal-ai/hunyuan3d-v3/image-to-3d

Transform your photos into ultra-high-resolution 3D models in seconds. Film-quality geometry with PBR textures, ready for games, e-commerce, and 3D printing.

text-to-speech
falREVIEW REQUIRED

Chatterbox

fal-ai/chatterbox/text-to-speech

Whether you're working on memes, videos, games, or AI agents, Chatterbox brings your content to life. Use the first tts from resemble ai.

text-to-speech
video-to-video
falREVIEW REQUIRED

Workflow Utilities Auto Subtitle

fal-ai/workflow-utilities/auto-subtitle

Add automatic subtitles to videos

auto-subtitlecaptioning
unknown
openrouterREVIEW REQUIRED

OpenRouter [Audio]

openrouter/router/audio

Run any audio capable LLM with fal. Process audio files — transcription, analysis, understanding, understand— using Gemini (Google) models. Supports wav, mp3, aiff, aac, ogg, flac, m4a. Powered by OpenRouter.

video-to-video
PixVerseREVIEW REQUIRED

PixVerse Lipsync

fal-ai/pixverse/lipsync

Generate realistic lipsync animations from audio using advanced algorithms for high-quality synchronization with PixVerse Lipsync model

animationlip sync
text-to-video
KlingREVIEW REQUIRED

Kling Video V3 Standard Turbo Text to Video

fal-ai/kling-video/v3/turbo/standard/text-to-video

Kling 3.0 Turbo Standard is a fast, cost-efficient video generation model that turns text prompts directly into 720P video with native audio, optimized for rapid iteration and high-volume production

stylizedtransformlipsync
text-to-image
RecraftREVIEW REQUIRED

Recraft V4.1 Text to Vector

fal-ai/recraft/v4.1/text-to-vector

Recraft V4.1 Vector turns prompts into fully editable SVGs with structured layers and clean geometry. Built for logos, icons, and illustration systems, it produces artwork that goes straight from generation into Figma or Illustrator.

stylizedtransformtypography
image-to-video
GoogleREVIEW REQUIRED

Veo 3.1 Fast

fal-ai/veo3.1/fast/reference-to-video

Generate videos from reference images using Google's Veo 3.1 Fast

text-to-audio
cassetteaiREVIEW REQUIRED

Sound Effects Generator

cassetteai/sound-effects-generator

Create stunningly realistic sound effects in seconds - CassetteAI's Sound Effects Model generates high-quality SFX up to 30 seconds long in just 1 second of processing time

soundsfxsound-effectscassetteai
image-to-video
lightricksREVIEW REQUIRED

LTX 2.5 Image to Video Pro

lightricks/ltx-2.5/image-to-video/pro

LTX-2.5 is Lightricks' open-source audio-video model. This endpoint animates a still image into video with synchronized audio in a single pass, in a quality-optimized mode for high-fidelity final output.

stylizedtransformlip-sync
image-to-video
KlingREVIEW REQUIRED

Kling Video

fal-ai/kling-video/o3/4k/image-to-video

Kling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling

stylizedtransformlipsync
video-to-video
falREVIEW REQUIRED

Workflow Utilities Scale Video

fal-ai/workflow-utilities/scale-video

FFMPEG Utilities to Scale Videos

image-to-video
MiniMaxREVIEW REQUIRED

MiniMax Hailuo 2.3 [Pro] (Image to Video)

fal-ai/minimax/hailuo-2.3/pro/image-to-video

MiniMax Hailuo-2.3 Image To Video API (Pro, 1080p): Advanced image-to-video generation model with 1080p resolution

image-to-video
image-to-image
Black Forest LabsREVIEW REQUIRED

Flux Pro Erase

fal-ai/flux-pro/v1/erase

Latest object erasing model from Black Forest Labs. Remove undesired objects, texts from images.

utilityediting
text-to-video
KlingREVIEW REQUIRED

Kling O3 Text to Video [Standard]

fal-ai/kling-video/o3/standard/text-to-video

Generate realistic videos using Kling O3 from Kling Team!

text-to-video
audio-to-audio
ElevenLabsREVIEW REQUIRED

ElevenLabs Audio Isolation

fal-ai/elevenlabs/audio-isolation

Isolate audio tracks using ElevenLabs advanced audio isolation technology.

audio
text-to-video
AlibabaREVIEW REQUIRED

Wan Text to Video

fal-ai/wan/v2.7/text-to-video

Wan 2.7 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.

stylizedtransformlipsync
image-to-3d
tripo3dREVIEW REQUIRED

Tripo3D

tripo3d/tripo/v2.5/image-to-3d

State of the art Image to 3D Object generation. Generate 3D model from a single image!

image-to-3dstylized