HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

206,822 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

any-to-any

google/gemma-4-12B-it-qat-q4_0-gguf

google/gemma-4-12B-it-qat-q4_0-gguf is a any to any model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

775,846289
12.0B12 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

bartowski/XYZAILab_XYZ-Aquila-mini-GGUF

bartowski/XYZAILab_XYZ-Aquila-mini-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

767,7834
Unknown8 GB+ VRAMapache-2.0
Deployment details
any-to-any

ggml-org/gemma-4-E4B-it-GGUF

ggml-org/gemma-4-E4B-it-GGUF is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

759,92484
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

cyankiwi/Qwen3.6-35B-A3B-AWQ-4bit

cyankiwi/Qwen3.6-35B-A3B-AWQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

759,55096
35.0B32 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

nvidia/parakeet-tdt-0.6b-v3

nvidia/parakeet-tdt-0.6b-v3 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

758,9371,107
600M4 GB+ VRAMcc-by-4.0
Deployment details
general AI

Comfy-Org/MiniMax-Music-3

Comfy-Org/MiniMax-Music-3 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

749,529225
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

Ransaka/sinlib

Ransaka/sinlib is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

742,8570
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

RedHatAI/gemma-4-31B-it-FP8-block

RedHatAI/gemma-4-31B-it-FP8-block is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

737,23445
31.0B80 GB+ VRAMapache-2.0
Deployment details
text-generation

FINAL-Bench/POCKET-35B-GGUF

FINAL-Bench/POCKET-35B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

733,56072
35.0B32 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

cyankiwi/Qwen3.5-4B-AWQ-4bit

cyankiwi/Qwen3.5-4B-AWQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

724,56721
4.0B6 GB+ VRAMapache-2.0
Deployment details
text-generation

openbmb/MiniCPM5-1B

openbmb/MiniCPM5-1B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

715,8691,105
1.0B6 GB+ VRAMapache-2.0
Deployment details
text-generation

microsoft/phi-4

microsoft/phi-4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

714,2002,297
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

unsloth/gemma-4-26B-A4B-it-qat-GGUF

unsloth/gemma-4-26B-A4B-it-qat-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

698,757409
26.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

RedHatAI/gemma-4-26B-A4B-it-FP8-dynamic

RedHatAI/gemma-4-26B-A4B-it-FP8-dynamic is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

694,95241
26.0B64 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

google/gemma-4-31B

google/gemma-4-31B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

681,874523
31.0B80 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/gemma-4-26B-A4B-it-GGUF

unsloth/gemma-4-26B-A4B-it-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

680,5801,108
26.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

ornith-ai/Ornith-1.5-35B-A3B-NVFP4

ornith-ai/Ornith-1.5-35B-A3B-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

674,81346
35.0B80 GB+ VRAMmit
Deployment details
image-text-to-text

tencent/HunyuanOCR

tencent/HunyuanOCR is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

666,606815
Unknown8 GB+ VRAMother
Deployment details
image-text-to-text

unsloth/gemma-4-31B-it-qat-GGUF

unsloth/gemma-4-31B-it-qat-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

665,647189
31.0B24 GB+ VRAMapache-2.0
Deployment details
general AI

kernels-community/flash-attn3

kernels-community/flash-attn3 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

659,28449
Unknown8 GB+ VRAMbsd-3-clause
Deployment details
image-text-to-text

unsloth/inkling-GGUF

unsloth/inkling-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

657,216135
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

poolside/Laguna-S-2.1-GGUF

poolside/Laguna-S-2.1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

655,173172
Unknown8 GB+ VRAMLicense unknown
Deployment details
image-text-to-text

zai-org/GLM-5.3-Flash

zai-org/GLM-5.3-Flash is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

654,9572,044
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

meta-models/Muse-Glimmer-30B

meta-models/Muse-Glimmer-30B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

648,7081,856
30.0B80 GB+ VRAMapache-2.0
Deployment details