text-to-imageSana
fal-ai/sanaSana can synthesize high-resolution, high-quality images with strong text-image alignment at a remarkably fast speed, with the ability to generate 4K images in less than a second.
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
text-to-imagefal-ai/sanaSana can synthesize high-resolution, high-quality images with strong text-image alignment at a remarkably fast speed, with the ability to generate 4K images in less than a second.
text-to-videofal-ai/kling-video/v3/4k/text-to-videoKling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling
image-to-3dfal-ai/hyper3d/rodin/v2Rodin by Hyper3D generates realistic and production ready 3D models from text or images.
image-to-3dfal-ai/meshy/v6-preview/image-to-3dMeshy-6-Preview is the latest model from Meshy. It generates realistic and production ready 3D models.
video-to-videofal-ai/ltx-2.3/reframeLTX-2.3 Reframe converts your videos to any aspect ratio without destructive cropping. It intelligently recenters the original footage and generatively fills the newly exposed areas with content that seamlessly matches the scene, so the result looks like it was shot natively in the target format. Turn landscape footage into vertical 9:16 for social, square 1:1 for feeds, or anything in between. Supports videos up to 60 seconds, with 720p and 1080p outputs across 1:1, 4:5, 5:4, 9:16 and 16:9.
text-to-audiofal-ai/mmaudio-v2/text-to-audioMMAudio generates synchronized audio given text inputs. It can generate sounds described by a prompt.
fal-ai/video-upscalerThe video upscaler endpoint uses RealESRGAN on each frame of the input video to upscale the video to a higher resolution.
text-to-imagerecraft/v4/style/text-to-imageGenerates raster images that hold a consistent style, from either a saved style ID or reference images attached directly.
video-to-videofal-ai/rife/videoInterpolate videos with RIFE - Real-Time Intermediate Flow Estimation
image-to-videofal-ai/pixverse/v5/image-to-videoGenerate high quality video clips from text and image prompts using PixVerse v5
audio-to-audiofal-ai/deepfilternet3Enhance speech audio by removing background noise and upsampling to 48KHz
video-to-audiosonilo/v1.1/video-to-musicAnalyzes your video’s pacing, mood, and timing to generate a frame-synced, licensed, commercial-use-safe soundtrack in seconds.
text-to-imagefal-ai/loraRun Any Stable Diffusion model with customizable LoRA weights.
image-to-3dhitem3d/hi3d/v3.0/image-to-3dGenerate 3D models from a single image with Hi3D V3.0.
text-to-audiofal-ai/stable-audio-3/small/music/text-to-audioStable Audio 3 Small Music is a 459 million parameter latent diffusion model that generates full stereo music compositions up to 2 minutes from text prompts, lightweight enough for on-device deployment.
image-to-imagefal-ai/phota/editPhota's model enables personalized photo editing, preserving identity while erasing distractions seamlessly.
text-to-video
text-to-imagefal-ai/ideogram/v3/generate-transparentGenerate images with transparent backgrounds using Ideogram Transparent model
text-to-videofal-ai/ltx-videoGenerate videos from prompts using LTX Video
image-to-imagemicrosoft/mai-image-2.5-pro/editApply precise, controllable edits to a reference image while preserving composition, typography, identity, and fine visual detail.
video-to-videominimax/h3/reference-to-video/loraReferences into video with synchronized audio using MiniMax H3
image-to-3dfal-ai/hyper3d/rodin/v2.5/fastRodin V2.5 by Hyper3D generates realistic and production ready 3D models from text or images. Do fast prototyping using the fast model.
image-to-imagefal-ai/flux-2/klein/4b/base/editImage-to-image editing with FLUX.2 [klein] 4B Base from Black Forest Labs. Precise modifications using natural language descriptions and hex color control.
video-to-videobria/video/background-removalAutomatically remove backgrounds from videos -perfect for creating clean, professional content without a green screen.
image-to-imagefal-ai/flux-2-lora-gallery/apartment-stagingVirtually furnishes an empty apartment
image-to-videofal-ai/minimax/hailuo-2.3-fast/pro/image-to-videoMiniMax Hailuo-2.3-Fast Image To Video API (Pro, 1080p): Advanced fast image-to-video generation model with 1080p resolution
image-to-3dfal-ai/trellis/multiGenerate 3D models from multiple images using Trellis. A native 3D generative model enabling versatile and high-quality 3D asset creation.
llmopenrouter/router/openai/v1/embeddingsGenerate text embeddings using OpenAI-compatible API. Access embedding models like text-embedding-3-small, text-embedding-3-large (OpenAI), and other embedding models available through OpenRouter. Drop-in replacement for the OpenAI embeddings API. Powered by OpenRouter.