EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 15 · 28 per page
text-to-image
microsoftREVIEW REQUIRED

Mai Image 2.5 Text to Image

microsoft/mai-image-2.5

MAI-Image-2.5 is Microsoft's photorealistic image generation and editing model that turns text prompts or uploaded images into high-quality, design-ready visuals with fine-grained, pixel-level control.

realismtypographystylized
3d-to-3d
MeshyREVIEW REQUIRED

Meshy Rigging Multi Animation

fal-ai/meshy/rigging/multi-animation

Meshy auto-rigs a humanoid 3D model fitting a skeleton and binding the mesh, then applies several motion presets from its animation library

stylizedtransform3D
video-to-video
veedREVIEW REQUIRED

Video Background Removal

veed/video-background-removal/fast

Remove background from any video with people and objects. No green screen needed.

text-to-image
RecraftREVIEW REQUIRED

Recraft V4 (Vector)

fal-ai/recraft/v4/text-to-vector

Recraft V4 was developed with designers to bring true visual taste to AI image generation. Built for brand systems and production-ready workflows, it goes beyond prompt accuracy — delivering stronger composition, refined lighting, realistic materials, and a cohesive aesthetic. The result is imagery shaped by professional design judgment, ready for immediate real-world use without additional post-processing.

text-to-imagetext-to-vector
text-to-speech
falREVIEW REQUIRED

Index TTS 2.0

fal-ai/index-tts-2/text-to-speech

Generate natural, clear speeches using Index TTS 2.0 from IndexTeam

text-to-speech
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev] Inpainting with LoRAs

fal-ai/flux-lora/inpainting

Super fast endpoint for the FLUX.1 [dev] inpainting model with LoRA support, enabling rapid and high-quality image inpaingting using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.

lorapersonalization
image-to-video
falREVIEW REQUIRED

Ffmpeg Api Images to Video

fal-ai/ffmpeg-api/images-to-video

A fal.ai endpoint that stitches an ordered list of images into an MP4 video by holding each image for a specified number of frames at a configurable frame rate

utilityediting
video-to-video
falREVIEW REQUIRED

Infinitalk

fal-ai/infinitalk

Infinitalk model generates a talking avatar video from an image and audio file. The avatar lip-syncs to the provided audio with natural facial expressions.

stylizedtransform
image-to-image
falREVIEW REQUIRED

Sam 3

fal-ai/sam-3/image-rle

SAM 3 is a unified foundation model for promptable segmentation in images and videos. It can detect, segment, and track objects using text or visual prompts such as points, boxes, and masks.

segmentationrlereal-time
image-to-image
falREVIEW REQUIRED

Midas Depth Estimation

fal-ai/imageutils/depth

Create depth maps using Midas depth estimation.

depthutility
text-to-image
OpenAIREVIEW REQUIRED

GPT Image 1 Mini

fal-ai/gpt-image-1-mini

GPT Image 1 mini combines OpenAI's advanced language capabilities, powered by GPT-5, with GPT Image 1 Mini for efficient image generation.

text-to-image
image-to-video
PixVerseREVIEW REQUIRED

PixVerse Swap

fal-ai/pixverse/swap

Generate high quality video clips by swapping person, objects and background using Pixverse Swap.

video-to-video
falREVIEW REQUIRED

Sam 3

fal-ai/sam-3/video

SAM 3 is a unified foundation model for promptable segmentation in images and videos. It can detect, segment, and track objects using text or visual prompts such as points, boxes, and masks.

segmentationmaskreal-time
speech-to-speech
falREVIEW REQUIRED

Chatterbox

fal-ai/chatterbox/speech-to-speech

Whether you're working on memes, videos, games, or AI agents, Chatterbox brings your content to life. Use the first tts from resemble ai.

speech-to-speech
image-to-image
topazREVIEW REQUIRED

Topaz Upscale Image Transparent

topaz/upscale/image/transparent

Professional transparent-image upscaling powered by Topaz Labs. Preserves the alpha channel end to end with PNG output. Best for logos, stickers and assets with transparency.

upscaleimage
image-to-video
KlingREVIEW REQUIRED

Kling O1 Reference Image to Video [Pro]

fal-ai/kling-video/o1/reference-to-video

Transform images, elements, and text into consistent, high-quality video scenes, ensuring stable character identity, object details, and environments.

json
falREVIEW REQUIRED

Ffmpeg Api

fal-ai/ffmpeg-api/loudnorm

Get EBU R128 loudness normalization from audio files using FFmpeg API.

ffmpeg
image-to-image
falREVIEW REQUIRED

PATINA

fal-ai/patina

PATINA creates seamless high-resolution normal, roughness, basecolor (albedo), height (displacement) and metalness maps from images

pbrdisplacementmetalnessnormal
text-to-3d
falREVIEW REQUIRED

Hunyuan 3D Pro Text to 3D

fal-ai/hunyuan-3d/v3.1/pro/text-to-3d

Generate 3D models from text prompts with Hunyuan 3D Pro

3dhunyuantext-to-3d
image-to-image
briaREVIEW REQUIRED

Bria Increase Resolution: Upscale Images up to 4x Without Losing Detail | fal

bria/increase-resolution

Upscale any image 2x or 4x, up to 8192×8192, with Bria Increase Resolution. Preserves the original content — no regeneration, no altered details. Commercial-safe

utilityediting
image-to-image
briaREVIEW REQUIRED

Fibo Edit 1.5 Image Editing

bria/fibo-edit-1.5/edit

Commercially safe, multi-reference image editing model. Follows natural language instructions alone or with up to 4 reference images, purpose-built for complex object and character combinations, virtual try-on, background replacement, style transfer, and more.

stylizedtransformediting
text-to-video
MiniMaxREVIEW REQUIRED

MiniMax Hailuo 2.3 [Standard] (Text to Video)

fal-ai/minimax/hailuo-2.3/standard/text-to-video

MiniMax Hailuo-2.3 Text To Video API (Standard, 768p): Advanced text-to-video generation model with 768p resolution

text-to-video
video-to-video
KlingREVIEW REQUIRED

Kling Video 4K Video to Video

fal-ai/kling-video/o3/4k/video-to-video/reference

Kling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling

utilityediting
image-to-image
falREVIEW REQUIRED

Image Editing Photo Restoration

fal-ai/image-editing/photo-restoration

Restore and enhance old or damaged photos by removing imperfections, adding color while preserving the original character and details of the image.

stylizedtransform
image-to-video
AlibabaREVIEW REQUIRED

Wan 2.7 Reference to Video

fal-ai/wan/v2.7/reference-to-video

Wan 2.7 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.

stylizedtransformlipsync
text-to-image
IdeogramREVIEW REQUIRED

V4.0q [fast]

ideogram/v4/fast

Generate high-quality images, posters, and logos with Ideogram's latest V4.0q — producing crisp visuals with accurate text rendering, fine detail, and full creative control for polished, ready-to-use designs IN A SECOND.

realismtypographystylized
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev] with Controlnets and Loras

fal-ai/flux-general/inpainting

FLUX General Inpainting is a versatile endpoint that enables precise image editing and completion, supporting multiple AI extensions including LoRA, ControlNet, and IP-Adapter for enhanced control over inpainting results and sophisticated image modifications.

loracontrolnetip-adapter