speech-to-textWizper (Whisper v3 -- fal.ai edition)
fal-ai/wizper[Experimental] Whisper v3 Large -- but optimized by our inference wizards. Same WER, double the performance!
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
speech-to-textfal-ai/wizper[Experimental] Whisper v3 Large -- but optimized by our inference wizards. Same WER, double the performance!
speech-to-textfal-ai/elevenlabs/speech-to-text/scribe-v2Use Scribe-V2 from ElevenLabs to do blazingly fast speech to text inferences!
fal-ai/elevenlabs/speech-to-textGenerate text from speech using ElevenLabs advanced speech-to-text model.
speech-to-textfal-ai/elevenlabs/forced-alignmentAlign the transcript and your audio recording using Elevenlab's forced alignment feature!
speech-to-textfal-ai/speech-to-textLeverage the rapid processing capabilities of AI models to enable accurate and efficient real-time speech-to-text transcription.
speech-to-textfal-ai/cohere-transcribeCohere Transcribe turns your business audio into accurate text, ready for search, analytics, and automation
speech-to-textnvidia/nemotron-asr-multilingual/asrNemotron-ASR-Streaming is a multi lingual, streaming Automatic Speech Recognition (ASR) engineered to deliver high-quality multi lingual transcription across both low-latency streaming and high-throughput batch workloads.
speech-to-textfal-ai/speech-to-text/turboLeverage the rapid processing capabilities of AI models to enable accurate and efficient real-time speech-to-text transcription.
speech-to-textfal-ai/smart-turnAn open source, community-driven and native audio turn detection model by Pipecat AI.
speech-to-textfal-ai/speech-to-text/streamLeverage the rapid processing capabilities of AI models to enable accurate and efficient real-time speech-to-text transcription.
speech-to-textfal-ai/speech-to-text/turbo/streamLeverage the rapid processing capabilities of AI models to enable accurate and efficient real-time speech-to-text transcription.