fixie-ai/ultravox-v0_5-llama-3_2-1b
fixie-ai/ultravox-v0_5-llama-3_2-1b is a audio text to text model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
What to verify before deployment
- Confirm the exact weight format and quantization.
- Measure memory at your intended context or image size.
- Review model-card limitations and evaluation methodology.
- Test latency and throughput on your target runtime.
Likely commercial-friendly
License metadata is an index signal, not legal advice. Follow the repository license and any model-specific acceptable-use terms.
Verify at the source →More audio-text-to-text models
fixie-ai/ultravox-v0_5-llama-3_2-1b
fixie-ai/ultravox-v0_5-llama-3_2-1b is a audio text to text model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
cstr/MOSS-Audio-4B-Instruct-GGUF
cstr/MOSS-Audio-4B-Instruct-GGUF is a audio text to text model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
mispeech/midashenglm-0.6b-fp32
mispeech/midashenglm-0.6b-fp32 is a audio text to text model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mispeech/midashenglm-0.6b-gguf
mispeech/midashenglm-0.6b-gguf is a audio text to text model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
snkii/Sori-1B
snkii/Sori-1B is a audio text to text model indexed for deployment research. Estimated minimum GPU memory is 6 GB. Access approval is required on Hugging Face.
Luigi/Qwen3-ASR-0.6B-Agent
Luigi/Qwen3-ASR-0.6B-Agent is a audio text to text model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
Operate a GPU cloud or inference API?
Reach developers after they have selected a model and are ready to run it.