HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

204,822 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

FINAL-Bench/POCKET-35B-GGUF

FINAL-Bench/POCKET-35B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

733,56072
35.0B32 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

cyankiwi/Qwen3.5-4B-AWQ-4bit

cyankiwi/Qwen3.5-4B-AWQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

724,56721
4.0B6 GB+ VRAMapache-2.0
Deployment details
text-generation

openbmb/MiniCPM5-1B

openbmb/MiniCPM5-1B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

715,8691,105
1.0B6 GB+ VRAMapache-2.0
Deployment details
text-generation

microsoft/phi-4

microsoft/phi-4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

714,2002,297
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

unsloth/gemma-4-26B-A4B-it-qat-GGUF

unsloth/gemma-4-26B-A4B-it-qat-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

698,757409
26.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

RedHatAI/gemma-4-26B-A4B-it-FP8-dynamic

RedHatAI/gemma-4-26B-A4B-it-FP8-dynamic is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

694,95241
26.0B64 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

google/gemma-4-31B

google/gemma-4-31B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

681,874523
31.0B80 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/gemma-4-26B-A4B-it-GGUF

unsloth/gemma-4-26B-A4B-it-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

680,5801,108
26.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

ornith-ai/Ornith-1.5-35B-A3B-NVFP4

ornith-ai/Ornith-1.5-35B-A3B-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

674,81346
35.0B80 GB+ VRAMmit
Deployment details
image-text-to-text

unsloth/gemma-4-31B-it-qat-GGUF

unsloth/gemma-4-31B-it-qat-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

665,647189
31.0B24 GB+ VRAMapache-2.0
Deployment details
general AI

kernels-community/flash-attn3

kernels-community/flash-attn3 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

659,28449
Unknown8 GB+ VRAMbsd-3-clause
Deployment details
image-text-to-text

unsloth/inkling-GGUF

unsloth/inkling-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

657,216135
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

zai-org/GLM-5.3-Flash

zai-org/GLM-5.3-Flash is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

654,9572,044
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

meta-models/Muse-Glimmer-30B

meta-models/Muse-Glimmer-30B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

648,7081,856
30.0B80 GB+ VRAMapache-2.0
Deployment details
any-to-any

google/gemma-4-E4B-it-qat-q4_0-gguf

google/gemma-4-E4B-it-qat-q4_0-gguf is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

646,856133
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

empero-ai/Qwythos-9B-v2-GGUF

empero-ai/Qwythos-9B-v2-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

636,661266
9.0B8 GB+ VRAMapache-2.0
Deployment details
text-generation

prism-ml/Ternary-Bonsai-27B-gguf

prism-ml/Ternary-Bonsai-27B-gguf is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

636,3431,269
27.0B24 GB+ VRAMapache-2.0
Deployment details
text-ranking

zeroentropy/zerank-2-reranker

zeroentropy/zerank-2-reranker is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

634,780117
Unknown8 GB+ VRAMapache-2.0
Deployment details
any-to-any

unsloth/gemma-4-E4B-it-qat-GGUF

unsloth/gemma-4-E4B-it-qat-GGUF is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

632,463174
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-to-video

larryvrh/MiniMax-H3-Turbo-Lora

larryvrh/MiniMax-H3-Turbo-Lora is a text to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

621,924932
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

cyankiwi/MiniCPM-SALA-AWQ-8bit

cyankiwi/MiniCPM-SALA-AWQ-8bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

608,7430
Unknown8 GB+ VRAMapache-2.0
Deployment details