HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

204,822 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

image-text-to-text

unsloth/gemma-4-E2B-it-GGUF

unsloth/gemma-4-E2B-it-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

589,736302
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

AEON-7/Qwen3.6-35B-A3B-heretic-NVFP4

AEON-7/Qwen3.6-35B-A3B-heretic-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

583,59271
35.0B80 GB+ VRAMapache-2.0
Deployment details
any-to-any

unsloth/gemma-4-E2B-it-qat-GGUF

unsloth/gemma-4-E2B-it-qat-GGUF is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

580,58374
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

google/gemma-4-26B-A4B-it-qat-q4_0-gguf

google/gemma-4-26B-A4B-it-qat-q4_0-gguf is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

569,450164
26.0B24 GB+ VRAMapache-2.0
Deployment details
any-to-any

google/gemma-4-E4B

google/gemma-4-E4B is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

559,734413
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

swiss-ai/Apertus-8B-Instruct-2509

swiss-ai/Apertus-8B-Instruct-2509 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

546,825489
8.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/gemma-4-E4B-it-GGUF

unsloth/gemma-4-E4B-it-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

545,755611
Unknown8 GB+ VRAMapache-2.0
Deployment details
any-to-any

google/gemma-4-E2B-it-qat-q4_0-gguf

google/gemma-4-E2B-it-qat-q4_0-gguf is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

545,397122
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

openbmb/MiniCPM-V-4.6

openbmb/MiniCPM-V-4.6 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

521,5561,206
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

meta-models/Muse-Glimmer-30B-GGUF

meta-models/Muse-Glimmer-30B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

508,379327
30.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

empero-ai/Qwen3.8-4B-Distill-GGUF

empero-ai/Qwen3.8-4B-Distill-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

508,25599
4.0B6 GB+ VRAMapache-2.0
Deployment details
text-generation

empero-ai/Qwen3.8-2B-Distill-GGUF

empero-ai/Qwen3.8-2B-Distill-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

503,511126
2.0B4 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

google/gemma-4-31B-it-qat-q4_0-gguf

google/gemma-4-31B-it-qat-q4_0-gguf is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

484,616127
31.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

cyankiwi/Qwen3-VL-8B-Instruct-AWQ-4bit

cyankiwi/Qwen3-VL-8B-Instruct-AWQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

481,56617
8.0B8 GB+ VRAMapache-2.0
Deployment details
text-generation

LilaRest/gemma-4-31B-it-NVFP4-turbo

LilaRest/gemma-4-31B-it-NVFP4-turbo is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

473,694303
31.0B80 GB+ VRAMapache-2.0
Deployment details