HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

394,301 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

image-text-to-text

AtomicChat/gemma-4-12B-it-GGUF

AtomicChat/gemma-4-12B-it-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

5,0566
12.0B12 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

mixbits/Qwen3.8-27B-NVFP4-MTP-VL-GGUF

mixbits/Qwen3.8-27B-NVFP4-MTP-VL-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

5,0557
27.0B24 GB+ VRAMapache-2.0
Deployment details
general AI

mudler/GLM-4.7-Flash-APEX-GGUF

mudler/GLM-4.7-Flash-APEX-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,05117
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Hcompany/Holo-3.1-0.8B

Hcompany/Holo-3.1-0.8B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

5,03228
800M4 GB+ VRAMapache-2.0
Deployment details
general AI

tencent/Hy-MT2-1.8B-1.25Bit-GGUF

tencent/Hy-MT2-1.8B-1.25Bit-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

5,02745
1.8B4 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

leok7v/Qwen3.8-27B

leok7v/Qwen3.8-27B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

5,0210
27.0B24 GB+ VRAMapache-2.0
Deployment details
general AI

enginetown/Qwen3.8-27B-Calibrated

enginetown/Qwen3.8-27B-Calibrated is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

5,01220
27.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

mlx-community/Qwen3.8-27B-MTP-mxfp8

mlx-community/Qwen3.8-27B-MTP-mxfp8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

5,0006
27.0B64 GB+ VRAMapache-2.0
Deployment details
text-generation

NC-AI-consortium-VAETKI/VAETKI

NC-AI-consortium-VAETKI/VAETKI is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,98662
Unknown8 GB+ VRAMmit
Deployment details
general AI

allenai/tmax-9b

allenai/tmax-9b is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

4,98216
9.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

mlx-community/Muse-Glimmer-30B-8bit

mlx-community/Muse-Glimmer-30B-8bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.

4,9613
30.0B48 GB+ VRAMapache-2.0
Deployment details
text-to-speech

Sunbird/orpheus-3b-tts-multilingual

Sunbird/orpheus-3b-tts-multilingual is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

4,9543
3.0B12 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

CohereLabs/command-a-plus-05-2026-w4a4

CohereLabs/command-a-plus-05-2026-w4a4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,928241
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

furiosa-ai/Qwen3-VL-2B-Thinking

furiosa-ai/Qwen3-VL-2B-Thinking is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,9190
2.0B8 GB+ VRAMapache-2.0
Deployment details