HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

396,301 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

image-text-to-text

furiosa-ai/Qwen3-VL-2B-Thinking

furiosa-ai/Qwen3-VL-2B-Thinking is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,9190
2.0B8 GB+ VRAMapache-2.0
Deployment details
text-generation

cyankiwi/Ornith-1.5-9B-AWQ-INT4

cyankiwi/Ornith-1.5-9B-AWQ-INT4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,9161
9.0B8 GB+ VRAMmit
Deployment details
visual-document-retrieval

vidore/colSmol-256M

vidore/colSmol-256M is a visual document retrieval model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,90421
Unknown8 GB+ VRAMmit
Deployment details
general AI

kernels-community/mamba-ssm

kernels-community/mamba-ssm is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,9044
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-video

vantagewithai/Bernini-R-GGUF-ComfyUI

vantagewithai/Bernini-R-GGUF-ComfyUI is a image text to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,88013
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

reaperdoesntknow/TopologicalQwen

reaperdoesntknow/TopologicalQwen is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,8721
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

llm-jp/llm-jp-4-33b-thinking-gguf

llm-jp/llm-jp-4-33b-thinking-gguf is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

4,86911
33.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Vontra/GLM-5.3-Flash-MLX-4bit-MTP

Vontra/GLM-5.3-Flash-MLX-4bit-MTP is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,8417
Unknown8 GB+ VRAMmit
Deployment details