HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

206,822 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

any-to-any

google/gemma-4-E4B-it-qat-q4_0-gguf

google/gemma-4-E4B-it-qat-q4_0-gguf is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

646,856133
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

empero-ai/Qwythos-9B-v2-GGUF

empero-ai/Qwythos-9B-v2-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

636,661266
9.0B8 GB+ VRAMapache-2.0
Deployment details
text-generation

prism-ml/Ternary-Bonsai-27B-gguf

prism-ml/Ternary-Bonsai-27B-gguf is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

636,3431,269
27.0B24 GB+ VRAMapache-2.0
Deployment details
text-ranking

zeroentropy/zerank-2-reranker

zeroentropy/zerank-2-reranker is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

634,780117
Unknown8 GB+ VRAMapache-2.0
Deployment details
any-to-any

unsloth/gemma-4-E4B-it-qat-GGUF

unsloth/gemma-4-E4B-it-qat-GGUF is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

632,463174
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-to-video

larryvrh/MiniMax-H3-Turbo-Lora

larryvrh/MiniMax-H3-Turbo-Lora is a text to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

621,924932
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-video

unsloth/MiniMax-H3-GGUF

unsloth/MiniMax-H3-GGUF is a image text to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

620,661261
Unknown8 GB+ VRAMother
Deployment details
text-generation

cyankiwi/MiniCPM-SALA-AWQ-8bit

cyankiwi/MiniCPM-SALA-AWQ-8bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

608,7430
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/gemma-4-E2B-it-GGUF

unsloth/gemma-4-E2B-it-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

589,736302
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

AEON-7/Qwen3.6-35B-A3B-heretic-NVFP4

AEON-7/Qwen3.6-35B-A3B-heretic-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

583,59271
35.0B80 GB+ VRAMapache-2.0
Deployment details
any-to-any

unsloth/gemma-4-E2B-it-qat-GGUF

unsloth/gemma-4-E2B-it-qat-GGUF is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

580,58374
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

google/gemma-4-26B-A4B-it-qat-q4_0-gguf

google/gemma-4-26B-A4B-it-qat-q4_0-gguf is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

569,450164
26.0B24 GB+ VRAMapache-2.0
Deployment details
any-to-any

google/gemma-4-E4B

google/gemma-4-E4B is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

559,734413
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

swiss-ai/Apertus-8B-Instruct-2509

swiss-ai/Apertus-8B-Instruct-2509 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

546,825489
8.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/gemma-4-E4B-it-GGUF

unsloth/gemma-4-E4B-it-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

545,755611
Unknown8 GB+ VRAMapache-2.0
Deployment details
any-to-any

google/gemma-4-E2B-it-qat-q4_0-gguf

google/gemma-4-E2B-it-qat-q4_0-gguf is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

545,397122
Unknown8 GB+ VRAMapache-2.0
Deployment details