HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

276,508 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

mlx-community/Qwen3.8-27B-MTP-8bit

mlx-community/Qwen3.8-27B-MTP-8bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

13,16517
27.0B32 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

pearsonkyle/Qwen3.8-27B-GPTQ-W4A16

pearsonkyle/Qwen3.8-27B-GPTQ-W4A16 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

13,0965
27.0B24 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

microsoft/VibeVoice-ASR-BitNet

microsoft/VibeVoice-ASR-BitNet is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

13,087195
Unknown8 GB+ VRAMmit
Deployment details
text-generation

bartowski/granite-4.2-30b-GGUF

bartowski/granite-4.2-30b-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

13,0877
30.0B24 GB+ VRAMapache-2.0
Deployment details
image-to-text

zhiyuanyou/DeQA-Score-Mix3

zhiyuanyou/DeQA-Score-Mix3 is a image to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

13,0406
Unknown8 GB+ VRAMmit
Deployment details
text-generation

bartowski/Ling-3.0-flash-GGUF

bartowski/Ling-3.0-flash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

13,01413
Unknown8 GB+ VRAMmit
Deployment details
text-generation

mlx-community/gpt-oss-20b-OptiQ-4bit

mlx-community/gpt-oss-20b-OptiQ-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 16 GB. It is publicly listed on Hugging Face.

12,9558
20.0B16 GB+ VRAMapache-2.0
Deployment details
text-to-speech

Audio8/Audio8-TTS-Preview-0.6b

Audio8/Audio8-TTS-Preview-0.6b is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

12,879388
600M4 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

cyankiwi/Qwen3.8-27B-AWQ-BF16-INT4

cyankiwi/Qwen3.8-27B-AWQ-BF16-INT4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

12,80116
27.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

mlx-community/Qwen3.8-27B-OptiQ-4bit

mlx-community/Qwen3.8-27B-OptiQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

12,75718
27.0B24 GB+ VRAMapache-2.0
Deployment details