HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

210,822 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

image-text-to-text

RedHatAI/gemma-4-31B-it-NVFP4

RedHatAI/gemma-4-31B-it-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

74,27956
31.0B80 GB+ VRAMapache-2.0
Deployment details
text-to-speech

nineninesix/gepard-1.0

nineninesix/gepard-1.0 is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

70,551132
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

AtomicChat/Qwen3.5-9B-DFlash-GGUF

AtomicChat/Qwen3.5-9B-DFlash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

70,4672
9.0B8 GB+ VRAMapache-2.0
Deployment details
any-to-any

google/gemma-4-E2B

google/gemma-4-E2B is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

70,372467
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

CrashOverrideX/Quillan-Ronin

CrashOverrideX/Quillan-Ronin is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

70,2713
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

AtomicChat/Qwen3-4B-DFlash-GGUF

AtomicChat/Qwen3-4B-DFlash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

69,9230
4.0B6 GB+ VRAMmit
Deployment details
text-generation

AtomicChat/Qwen3.5-4B-DFlash-GGUF

AtomicChat/Qwen3.5-4B-DFlash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

69,6120
4.0B6 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

QuantTrio/gemma-4-31B-it-AWQ

QuantTrio/gemma-4-31B-it-AWQ is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

69,08214
31.0B80 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/Qwen3.8-27B-FP8

unsloth/Qwen3.8-27B-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

68,91628
27.0B64 GB+ VRAMapache-2.0
Deployment details
general AI

BAAI/seggpt-vit-large

BAAI/seggpt-vit-large is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

68,6225
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

cyankiwi/Qwen3-VL-2B-Instruct-AWQ-4bit

cyankiwi/Qwen3-VL-2B-Instruct-AWQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

68,4631
2.0B4 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

twolven/Qwen3.8-27B-abliterated-AWQ-MTP

twolven/Qwen3.8-27B-abliterated-AWQ-MTP is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

67,56014
27.0B64 GB+ VRAMapache-2.0
Deployment details
image-text-to-textGated

orcarouter/Qwen3.8-27B-Uncensored

orcarouter/Qwen3.8-27B-Uncensored is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. Access approval is required on Hugging Face.

67,011232
27.0B64 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

openbmb/MiniCPM-V-4.6-Thinking-BNB

openbmb/MiniCPM-V-4.6-Thinking-BNB is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

66,7604
Unknown8 GB+ VRAMapache-2.0
Deployment details