Detected Skills:
Found 0 registries and 11 entities for "Speech Recognition"
All Agents & Models
@cf/openai/whisper
modelWhisper is a general-purpose speech recognition model. It is trained on a large dataset of diverse audio and is also a multitasking model that can perform multilingual speech recognition, speech translation, and language identification.
@cf/deepgram/flux
modelFlux is the first conversational speech recognition model built specifically for voice agents.
@cf/openai/whisper-tiny-en
modelWhisper is a pre-trained model for automatic speech recognition (ASR) and speech translation. Trained on 680k hours of labelled data, Whisper models demonstrate a strong ability to generalize to many datasets and domains without the need for fine-tuning. This is the English-only version of the Whisper Tiny model which was trained on the task of speech recognition.
@cf/openai/whisper-large-v3-turbo
modelWhisper is a pre-trained model for automatic speech recognition (ASR) and speech translation.
subformer/meta-omnilingual-asr-7b
modelOmnilingual ASR 7B by Meta (Unofficial) - Automatic speech recognition supporting 1,693 languages with best-in-class accuracy. Meta's recommended variant for mission-critical transcription requiring maximum quality.
ibm-granite/granite-speech-3.3-8b
modelGranite-speech-3.3-8b is a compact and efficient speech-language model, specifically designed for automatic speech recognition (ASR) and automatic speech translation (AST).
erium/whisperx
modelAutomatic Speech Recognition with Word-level Timestamps & Diarization
nateraw/whisper-large-v3
modelWhisper is a general-purpose speech recognition model.
awerks/whisperx
modelFast automatic speech recognition (70x realtime with large-v2) with word-level timestamps and speaker diarization.
twangodev/qwenasr
modelServe QwenASR speech recognition and alignment.
ibm-granite/granite-speech-4.1-2b
modelGranite Speech 4.1 2B is a compact and efficient speech-language model, specifically designed for multilingual automatic speech recognition (ASR) and bidirectional automatic speech translation (AST) for English, French, German, Spanish, Portuguese and Jap