HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

398,301 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

image-text-to-text

Vontra/GLM-5.3-Flash-MLX-4bit-MTP

Vontra/GLM-5.3-Flash-MLX-4bit-MTP is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,8417
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

Minachist/Qwen3.8-27B-INT8-AutoRound

Minachist/Qwen3.8-27B-INT8-AutoRound is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

4,82911
27.0B32 GB+ VRAMapache-2.0
Deployment details
fill-mask

Synthyra/ESMplusplus_large

Synthyra/ESMplusplus_large is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,82817
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

unsloth/gemma-4-E4B-it-UD-MLX-4bit

unsloth/gemma-4-E4B-it-UD-MLX-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,80745
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

furiosa-ai/Qwen3-VL-4B-Thinking

furiosa-ai/Qwen3-VL-4B-Thinking is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

4,7970
4.0B12 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Vishva007/Qwen3.8-27B-W4A16-AutoRound

Vishva007/Qwen3.8-27B-W4A16-AutoRound is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

4,7862
27.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

furiosa-ai/Qwen3-8B-FP8

furiosa-ai/Qwen3-8B-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

4,7840
8.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

byteshape/Qwen3.6-35B-A3B-GGUF

byteshape/Qwen3.6-35B-A3B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

4,77939
35.0B32 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

handy-computer/whisper-tiny-gguf

handy-computer/whisper-tiny-gguf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,7741
Unknown8 GB+ VRAMapache-2.0
Deployment details
translation

unsloth/Hy-MT2-7B-GGUF

unsloth/Hy-MT2-7B-GGUF is a translation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,77121
7.0B8 GB+ VRAMapache-2.0
Deployment details