microsoft/Phi-4-multimodal-instruct
mit🎉Phi-4: mini-reasoning( | reasoning( | multimodal-instruct( | onnx( mini-instruct( | onnx( Model Summary Phi-4-multimodal-instruct is a lig...
ggerganov/whisper.cpp
mitOpenAI's Whisper models converted to ggml format for use with whisper.cpp( Available models( | Model | Disk | SHA | | ------------------- | ...
pyannote/speaker-diarization-3.1
mitVisit HuggingFace for more details.
distil-whisper/distil-large-v2
mitDistil-Whisper: distil-large-v2 Distil-Whisper was proposed in the paper Robust Knowledge Distillation via Large-Scale Pseudo Labelling( It ...
openai/whisper-small
apache-2.0Whisper Whisper is a pre-trained model for automatic speech recognition (ASR) and speech translation. Trained on 680k hours of labelled data...
facebook/wav2vec2-base-960h
apache-2.0Wav2Vec2-Base-960h Facebook's Wav2Vec2( The base model pretrained and fine-tuned on 960 hours of Librispeech on 16kHz sampled speech audio. ...
nyrahealth/CrisperWhisper
cc-by-nc-4.0CrisperWhisper > ⚠️ Deprecation notice > > CrisperWhisper (v1) is superseded by CrisperWhisper 2.0( and is no longer actively maintained. Cr...
openai/whisper-base
apache-2.0Whisper Whisper is a pre-trained model for automatic speech recognition (ASR) and speech translation. Trained on 680k hours of labelled data...
pyannote/speaker-diarization-3.0
mitVisit HuggingFace for more details.
pyannote/speaker-diarization
mitVisit HuggingFace for more details.