HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

390,301 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

general AI

mradermacher/Holo-3.1-4B-GGUF

mradermacher/Holo-3.1-4B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

5,1747
4.0B6 GB+ VRAMapache-2.0
Deployment details
text-generation

mradermacher/Qwen3.8-Queen-27B-GGUF

mradermacher/Qwen3.8-Queen-27B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

5,1692
27.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

RedHatAI/GLM-5.2-MXFP4xFP8_BLOCK

RedHatAI/GLM-5.2-MXFP4xFP8_BLOCK is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,1660
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

stepfun-ai/Step-3.7-Flash-GGUF

stepfun-ai/Step-3.7-Flash-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,159172
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

reaperdoesntknow/TameForCasualLM

reaperdoesntknow/TameForCasualLM is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,1550
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

z-lab/Muse-Glimmer-30B-DFlash2-GGUF

z-lab/Muse-Glimmer-30B-DFlash2-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

5,13712
30.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/Qwen3.5-27B

unsloth/Qwen3.5-27B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

5,11416
27.0B64 GB+ VRAMapache-2.0
Deployment details
text-to-speechGated

neuphonic/neutts-air

neuphonic/neutts-air is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.

5,113889
Unknown8 GB+ VRAMapache-2.0
Deployment details