context.vn

Small Model Hub

3,378 efficient AI models across 7 domains · Charts · API

ScorePercentile rank within each task — comparable across domains
Showing 91 models
#ModelDomainTaskParams ↑GFLOPsVRAMContextSpeedArchSource
1
nvidia/stt_en_conformer_ctc_small
SpeechSpeech Recognition10.0M24 MB6258 RTFxConformer-based/CTCOpen ASR
2
nvidia/stt_en_conformer_transducer_small
SpeechSpeech Recognition14.0M34 MB5557 RTFxConformer-based/TDT / RNN-TOpen ASR
3
abr-ai/niagara-19m-batch.en
SpeechSpeech Recognition20.0M48 MB3735 RTFxOpen ASR
4
usefulsensors/moonshine-streaming-tiny
SpeechSpeech Recognition30.0M72 MB4375 RTFxCustom/TransformerOpen ASR
5
usefulsensors/moonshine-tiny
SpeechSpeech Recognition30.0M72 MB3733 RTFxCustom/TransformerOpen ASR
6
abr-ai/niagara-38m-batch.en
SpeechSpeech Recognition38.0M91 MB4049 RTFxOpen ASR
7
usefulsensors/moonshine-base
SpeechSpeech Recognition60.0M144 MB2767 RTFxCustom/TransformerOpen ASR
8
facebook/data2vec-audio-base-960h
SpeechSpeech Recognition90.0M216 MB2930 RTFxSelf-supervised/CTCOpen ASR
9
facebook/wav2vec2-base-960h
SpeechSpeech Recognition90.0M216 MB3366 RTFxSelf-supervised/CTCOpen ASR
10
nvidia/parakeet-tdt_ctc-110m
SpeechSpeech Recognition110.0M264 MB6119 RTFxConformer-based/"CTCOpen ASR
11
nvidia/stt_en_fastconformer_transducer_large
SpeechSpeech Recognition110.0M264 MB6033 RTFxConformer-based/TDT / RNN-TOpen ASR
12
nvidia/stt_en_conformer_ctc_large
SpeechSpeech Recognition120.0M288 MB5732 RTFxConformer-based/CTCOpen ASR
13
nvidia/stt_en_fastconformer_ctc_large
SpeechSpeech Recognition120.0M288 MB6027 RTFxConformer-based/CTCOpen ASR
14
usefulsensors/moonshine-streaming-small
SpeechSpeech Recognition120.0M288 MB3206 RTFxCustom/TransformerOpen ASR
15
nvidia/canary-180m-flash
SpeechSpeech Recognition180.0M432 MB2484 RTFxConformer-based/TransformerOpen ASR
16
usefulsensors/moonshine-streaming-medium
SpeechSpeech Recognition240.0M576 MB2681 RTFxCustom/TransformerOpen ASR
17
soundsgoodai/Zipformer-cr-ctc-transducer-XL-290M
SpeechSpeech Recognition290.0M696 MB160 RTFxZipformer/Pruned Stateless Transducer with k2 modified_beam_searchOpen ASR
18
soundsgoodai/Zipformer-transducer-XL-290M
SpeechSpeech Recognition290.0M696 MB141 RTFxZipformer/RNN-TOpen ASR
19
facebook/wav2vec2-large-robust-ft-libri-960h
SpeechSpeech Recognition300.0M720 MB3058 RTFxSelf-supervised/CTCOpen ASR
20
facebook/data2vec-audio-large-960h
SpeechSpeech Recognition310.0M744 MB2374 RTFxSelf-supervised/CTCOpen ASR
21
facebook/hubert-large-ls960-ft
SpeechSpeech Recognition320.0M768 MB3024 RTFxSelf-supervised/CTCOpen ASR
22
facebook/wav2vec2-large-960h
SpeechSpeech Recognition320.0M768 MB3614 RTFxSelf-supervised/CTCOpen ASR
23
facebook/wav2vec2-large-960h-lv60-self
SpeechSpeech Recognition320.0M768 MB3010 RTFxSelf-supervised/CTCOpen ASR
24
speechbrain/asr-wav2vec2-librispeech
SpeechSpeech Recognition320.0M768 MB3628 RTFxSelf-supervised/CTCOpen ASR
25
AutoArk-AI/Audio8-ASR-0.1B
SpeechSpeech Recognition324.0M778 MB709 RTFxQwen3-ASR audio encoder/8-layer Qwen-style causal LMOpen ASR
26
speechbrain/asr-conformer-largescaleasr
SpeechSpeech Recognition480.0M1.1 GB72 RTFxConformer-based/"TransformerOpen ASR
27
facebook/wav2vec2-conformer-rope-large-960h-ft
SpeechSpeech Recognition600.0M1.4 GB1723 RTFx"Self-supervised/Conformer-based"Open ASR
28
nvidia/parakeet-ctc-0.6b
SpeechSpeech Recognition600.0M1.4 GB5884 RTFxConformer-based/CTCOpen ASR
29
nvidia/parakeet-rnnt-0.6b
SpeechSpeech Recognition600.0M1.4 GB5407 RTFxConformer-based/TDT / RNN-TOpen ASR
30
nvidia/parakeet-tdt-0.6b-v2
SpeechSpeech Recognition600.0M1.4 GB6038 RTFxConformer-based/TDT / RNN-TOpen ASR
31
nvidia/parakeet-tdt-0.6b-v3
SpeechSpeech Recognition600.0M1.4 GB6098 RTFxConformer-based/TDT / RNN-TOpen ASR
32
facebook/wav2vec2-conformer-rel-pos-large-960h-ft
SpeechSpeech Recognition620.0M1.5 GB1415 RTFx"Self-supervised/Conformer-based"Open ASR
33
nvidia/nemotron-speech-streaming-en-0.6b
SpeechSpeech Recognition620.0M1.5 GB1071 RTFxFastConformer (cache-aware)/RNNTOpen ASR
34
nvidia/nemotron-3.5-asr-streaming-0.6b
SpeechSpeech Recognition640.0M1.5 GB1490 RTFxFastConformer (cache-aware)/RNNTOpen ASR
35
efficient-speech/lite-whisper-large-v3-turbo-acc
SpeechSpeech Recognition700.0M1.6 GB327 RTFxWhisper/TransformerOpen ASR
36
Qwen/Qwen3-ASR-0.6B
SpeechSpeech Recognition780.0M1.8 GB439 RTFxCustom/Qwen3Open ASR
37
Qwen/Qwen3-ASR-0.6B-hf
SpeechSpeech Recognition780.0M1.8 GB730 RTFxCustom/Qwen3Open ASR
38
distil-whisper/distil-large-v3.5
SpeechSpeech Recognition800.0M1.9 GB874 RTFxWhisper/TransformerOpen ASR
39
openai/whisper-large-v3-turbo
SpeechSpeech Recognition800.0M1.9 GB783 RTFxWhisper/TransformerOpen ASR
40
facebook/hubert-xlarge-ls960-ft
SpeechSpeech Recognition960.0M2.3 GB1980 RTFxSelf-supervised/CTCOpen ASR
41
facebook/mms-1b-all
SpeechSpeech Recognition960.0M2.3 GB1954 RTFxSelf-supervised/CTCOpen ASR
42
efficient-speech/lite-whisper-large-v3
SpeechSpeech Recognition1.0B2.3 GB212 RTFxWhisper/TransformerOpen ASR
43
efficient-speech/lite-whisper-large-v3-acc
SpeechSpeech Recognition1.0B2.3 GB204 RTFxWhisper/TransformerOpen ASR
44
efficient-speech/lite-whisper-large-v3-fast
SpeechSpeech Recognition1.0B2.3 GB212 RTFxWhisper/TransformerOpen ASR
45
espnet/owsm_ctc_v3.2_ft_1B
SpeechSpeech Recognition1.0B2.3 GB692 RTFxWhisper/CTCOpen ASR
46
espnet/owsm_ctc_v4_1B
SpeechSpeech Recognition1.0B2.3 GB765 RTFxWhisper/CTCOpen ASR
47
nvidia/canary-1b
SpeechSpeech Recognition1.0B2.3 GB766 RTFxConformer-based/TransformerOpen ASR
48
nvidia/canary-1b-flash
SpeechSpeech Recognition1.0B2.3 GB2126 RTFxConformer-based/TransformerOpen ASR
49
nvidia/canary-1b-v2
SpeechSpeech Recognition1.0B2.3 GB1821 RTFxConformer-based/TransformerOpen ASR
50
espnet/owsm_ctc_v3.1_1B
SpeechSpeech Recognition1.1B2.6 GB816 RTFxWhisper/CTCOpen ASR
51
nvidia/parakeet-ctc-1.1b
SpeechSpeech Recognition1.1B2.6 GB5015 RTFxConformer-based/CTCOpen ASR
52
nvidia/parakeet-rnnt-1.1b
SpeechSpeech Recognition1.1B2.6 GB4122 RTFxConformer-based/TDT / RNN-TOpen ASR
53
nvidia/parakeet-tdt-1.1b
SpeechSpeech Recognition1.1B2.6 GB4529 RTFxConformer-based/TDT / RNN-TOpen ASR
54
AutoArk-AI/ARK-ASR-0.6B
SpeechSpeech Recognition1.1B2.7 GB672 RTFxWhisper/Qwen2Open ASR
55
CohereLabs/cohere-transcribe-03-2026
SpeechSpeech Recognition2.0B4.7 GB916 RTFxConformer-based/TransformerOpen ASR
56
ibm-granite/granite-4.0-1b-speech
SpeechSpeech Recognition2.0B4.7 GB661 RTFxConformer-based/GraniteOpen ASR
57
ibm-granite/granite-speech-4.1-2b
SpeechSpeech Recognition2.0B4.7 GB547 RTFxConformer-based/GraniteOpen ASR
58
ibm-granite/granite-speech-4.1-2b-nar
SpeechSpeech Recognition2.0B4.7 GB2079 RTFxConformer-based/Hybrid CTC-LLMOpen ASR
59
nyrahealth/CrisperWhisper
SpeechSpeech Recognition2.0B4.7 GB33 RTFxWhisper/TransformerOpen ASR
60
openai/whisper-large-v3
SpeechSpeech Recognition2.0B4.7 GB462 RTFxWhisper/TransformerOpen ASR
61
zai-org/GLM-ASR-Nano-2512
SpeechSpeech Recognition2.0B4.7 GB333 RTFxWhisper/LlamaOpen ASR
62
Qwen/Qwen3-ASR-1.7B
SpeechSpeech Recognition2.0B4.8 GB394 RTFxCustom/Qwen3Open ASR
63
Qwen/Qwen3-ASR-1.7B-hf
SpeechSpeech Recognition2.0B4.8 GB796 RTFxCustom/Qwen3Open ASR
64
OpenMOSS-Team/MOSS-Transcribe-preview-2B
SpeechSpeech Recognition2.4B5.7 GB149 RTFxQwen3-Omni-MoE audio encoder/Qwen3Open ASR
65
nvidia/canary-qwen-2.5b
SpeechSpeech Recognition2.5B5.9 GB861 RTFxConformer-based/Qwen2Open ASR
66
kyutai/stt-2.6b-en
SpeechSpeech Recognition2.6B6.1 GB133 RTFxMimi tokenizer/TransformerOpen ASR
67
bosonai/higgs-audio-v3-stt
SpeechSpeech Recognition2.7B6.3 GB110 RTFxWhisper/Qwen3Open ASR
68
ibm-granite/granite-speech-3.3-2b
SpeechSpeech Recognition3.0B7.0 GB507 RTFxConformer-based/GraniteOpen ASR
69
AutoArk-AI/ARK-ASR-3B
SpeechSpeech Recognition3.8B8.8 GB484 RTFxWhisper/Qwen2Open ASR
70
mistralai/Voxtral-Mini-4B-Realtime-2602
SpeechSpeech Recognition4.0B9.4 GB105 RTFxCustom/TransformerOpen ASR
71
mistralai/Voxtral-Mini-3B-2507
SpeechSpeech Recognition5.0B11.7 GB180 RTFxWhisper/Ministral 3BOpen ASR
72
HojoAI/Hojo-ASR-V1
SpeechSpeech Recognition5.2B12.1 GB74 RTFxQwen3-Omni encoder/Qwen3-4BOpen ASR
73
microsoft/Phi-4-multimodal-instruct
SpeechSpeech Recognition6.0B14.1 GB163 RTFxConformer-based/Phi-4-Mini-InstructOpen ASR
74
facebook/omniASR-CTC-7B-v2
SpeechSpeech Recognition6.5B15.2 GB528 RTFxSelf-supervised/CTCOpen ASR
75
facebook/omniASR-LLM-7B-v2
SpeechSpeech Recognition7.8B18.3 GB143 RTFxSelf-supervised/TransformerOpen ASR
76
microsoft/VibeVoice-ASR-HF
SpeechSpeech Recognition8.0B18.8 GB221 RTFxCustom/Qwen2Open ASR
77
bosonai/higgs-audio-v3-8b-stt-v2
SpeechSpeech Recognition8.9B20.9 GB139 RTFxWhisper/Qwen3Open ASR
78
ibm-granite/granite-speech-3.3-8b
SpeechSpeech Recognition9.0B21.1 GB264 RTFxConformer-based/GraniteOpen ASR
79
mistralai/Voxtral-Small-24B-2507
SpeechSpeech Recognition24.0B56.3 GB100 RTFxWhisper/Ministral 24BOpen ASR
80
aquavoice/avalon-v1-en
SpeechSpeech RecognitionOpen ASR
81
assemblyai/universal-3-5-pro
SpeechSpeech RecognitionOpen ASR
82
assemblyai/universal-3-pro
SpeechSpeech RecognitionOpen ASR
83
elevenlabs/scribe_v2
SpeechSpeech RecognitionOpen ASR
84
gladia/solaria-3
SpeechSpeech RecognitionOpen ASR
85
microsoft/azure-speech-06-2026
SpeechSpeech RecognitionOpen ASR
86
modulate/vfast
SpeechSpeech RecognitionOpen ASR
87
reson8/resonant-1
SpeechSpeech RecognitionOpen ASR
88
reson8/resonant-1-flash
SpeechSpeech RecognitionOpen ASR
89
smallestai/pulse
SpeechSpeech RecognitionOpen ASR
90
speechmatics/enhanced
SpeechSpeech RecognitionOpen ASR
91
zoom/scribe_v1
SpeechSpeech RecognitionOpen ASR