video-to-videoPixelcut Video Background Removal
pixelcut/video-background-removalPixelcut's Video Background Remover is an AI segmentation model that erases backgrounds frame by frame, with seamless temporal consistency.
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
video-to-videopixelcut/video-background-removalPixelcut's Video Background Remover is an AI segmentation model that erases backgrounds frame by frame, with seamless temporal consistency.
image-to-videofal-ai/pixverse/v4.5/image-to-videoGenerate high quality video clips from text and image prompts using PixVerse v4.5
text-to-imagefal-ai/wan/v2.7/pro/text-to-imageGenerate premium-quality images from text prompts using the enhanced WAN 2.7 Pro model with superior detail and composition.
text-to-speechfal-ai/chatterbox/text-to-speech/multilingualWhether you're working on memes, videos, games, or AI agents, Chatterbox brings your content to life. Use the first tts from resemble ai.
image-to-video
audio-to-videofal-ai/elevenlabs/dubbingGenerate dubbed videos or audios using ElevenLabs Dubbing feature!
image-to-imagegoogle/virtual-try-onGenerate realistic virtual try-on images from a person image and a clothing product image.
image-to-imagefal-ai/bria/product-shotPlace any product in any scenery with just a prompt or reference image while maintaining high integrity of the product. Trained exclusively on licensed data for safe and risk-free commercial use and optimized for eCommerce.
image-to-imagefal-ai/image-apps-v2/photo-restorationRestore old or damaged photos by fixing colors, scratches, and resolution.
image-to-imagetopaz/restore/imageProfessional image restoration powered by Topaz Labs. Recover 3 generatively rebuilds natural detail; Dust-Scratch V2 cleans film dust and scratches. Best for old, damaged or degraded photos.
image-to-videofal-ai/pixverse/v5.5/image-to-videoGenerate high quality video clips from text and image prompts using PixVerse v5.5
text-to-imagefal-ai/qwen-image-2512/loraLoRA inference endpoint for Qwen Image 2512, an improved version of Qwen Image with better text rendering, finer natural textures, and more realistic human generation.
llmopenrouter/router/openai/v1/responsesThe OpenRouter Responses API with fal, powered by OpenRouter, provides unified access to a wide range of large language models - including GPT, Claude, Gemini, and many others through a single API interface.
video-to-videomirelo-ai/sfx-v1.5/video-to-videoGenerate synced sounds for any video, and return it with its new sound track (like MMAudio)
text-to-speechfal-ai/minimax/speech-2.6-turboGenerate speech from text prompts and different voices using the MiniMax Speech-2.6 HD model, which leverages advanced AI techniques to create high-quality text-to-speech.
image-to-textnvidia/nemotron-3-nano-omni/visionVision reasoning variant of NVIDIA's Nemotron 3 Nano Omni. 30B A3B hybrid Transformer-Mamba MoE - accepts an image plus a prompt and returns text.
text-to-videofal-ai/minimax/video-01Generate video clips from your prompts using MiniMax model
image-to-imageluma/agent/uni-1/v1/editLuma Uni-1 Edit reworks a source image from a text instruction, preserving the original composition while applying style changes and following optional reference images to steer the result.
text-to-imagefal-ai/flux-1/devFLUX.1 [dev] is a 12 billion parameter flow transformer that generates high-quality images from text. It is suitable for personal and commercial use.
image-to-videofal-ai/kling-video/v1/pro/ai-avatarKling AI Avatar Pro: The premium endpoint for creating avatar videos with realistic humans, animals, cartoons, or stylized characters
image-to-videoblackforestlabs/flux-3/keyframes-to-videoFLUX 3 is Black Forest Labs' frontier video model. This endpoint builds video from a sequence of keyframes, generating the motion between each anchor point for precise control over how a shot progresses.
image-to-imagefal-ai/sam-3-1/image-rleSAM 3.1 builds comes with Object Multiplex, a shared-memory approach for joint multi-object tracking that delivers faster speeds with larger number of objects tracked.
text-to-imagekrea/v2/medium/turbo/text-to-imageGenerate high-fidelity images extremely fast from text with Krea 2 Medium Turbo, supporting aspect ratio, creativity, seed controls, and optional style references.
image-to-imagefal-ai/hy-wu-editImage editing with HY-WU. Transfer outfits, swap faces, and blend textures instantly—no finetuning needed, just describe what you want and provide reference images.
video-to-videoxai/grok-imagine-video/extend-videoExtend videos with xAI's Grok Imagine video model
image-to-videofal-ai/ltx-video/image-to-videoGenerate videos from images using LTX Video
image-to-imagefal-ai/florence-2-large/ocr-with-regionFlorence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks
image-to-3dfal-ai/hunyuan3d/v2/multi-viewGenerate 3D models from your images using Hunyuan 3D. A native 3D generative model enabling versatile and high-quality 3D asset creation.