HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

268,509 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

general AI

kernels-community/cv-utils

kernels-community/cv-utils is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

19,0301
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

XHToken/Spark-X2.5-4B-GGUF

XHToken/Spark-X2.5-4B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

18,98248
4.0B6 GB+ VRAMapache-2.0
Deployment details
text-generation

shisa-ai/Ornith-1.0-35B-FP8-BLOCK

shisa-ai/Ornith-1.0-35B-FP8-BLOCK is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

18,8740
35.0B80 GB+ VRAMmit
Deployment details
image-text-to-text

RedHatAI/diffusiongemma-26B-A4B-it-NVFP4

RedHatAI/diffusiongemma-26B-A4B-it-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

18,87019
26.0B64 GB+ VRAMapache-2.0
Deployment details
text-generation

cyankiwi/Qwen3-Coder-Next-AWQ-4bit

cyankiwi/Qwen3-Coder-Next-AWQ-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

18,74231
Unknown8 GB+ VRAMapache-2.0
Deployment details
visual-document-retrieval

vidore/colpali-v1.3

vidore/colpali-v1.3 is a visual document retrieval model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

18,723100
Unknown8 GB+ VRAMmit
Deployment details
other

aurekaresearch/OpenDDE

aurekaresearch/OpenDDE is a other model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

18,2839
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

kernels-community/flash-attn2

kernels-community/flash-attn2 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

18,26033
Unknown8 GB+ VRAMbsd-3-clause
Deployment details
text-to-speech

OpenMOSS-Team/MOSS-TTS-Realtime

OpenMOSS-Team/MOSS-TTS-Realtime is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

18,028104
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

ggml-org/gemma-4-31B-it-GGUF

ggml-org/gemma-4-31B-it-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

17,98144
31.0B24 GB+ VRAMapache-2.0
Deployment details