UrocyonF/Qwen3-ASR-1.7B-NVFP4
UrocyonF/Qwen3-ASR-1.7B-NVFP4 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
What to verify before deployment
- Confirm the exact weight format and quantization.
- Measure memory at your intended context or image size.
- Review model-card limitations and evaluation methodology.
- Test latency and throughput on your target runtime.
Likely commercial-friendly
License metadata is an index signal, not legal advice. Follow the repository license and any model-specific acceptable-use terms.
Verify at the source →More automatic-speech-recognition models
Qwen/Qwen3-ASR-1.7B
Qwen/Qwen3-ASR-1.7B is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
mlx-community/parakeet-tdt-0.6b-v3
mlx-community/parakeet-tdt-0.6b-v3 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
handy-computer/nemotron-3.5-asr-streaming-0.6b-gguf
handy-computer/nemotron-3.5-asr-streaming-0.6b-gguf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mlx-community/parakeet-tdt-0.6b-v2
mlx-community/parakeet-tdt-0.6b-v2 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
handy-computer/parakeet-unified-en-0.6b-gguf
handy-computer/parakeet-unified-en-0.6b-gguf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
NbAiLab/nb-wav2vec2-1b-nynorsk
NbAiLab/nb-wav2vec2-1b-nynorsk is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
Operate a GPU cloud or inference API?
Reach developers after they have selected a model and are ready to run it.