HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

268,509 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

general AI

mudler/Ornith-1.5-35B-A3B-APEX-GGUF

mudler/Ornith-1.5-35B-A3B-APEX-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

17,80529
35.0B32 GB+ VRAMapache-2.0
Deployment details
general AI

gbuzhf/KAT-Coder-V2.5-Dev-MTP-GGUF

gbuzhf/KAT-Coder-V2.5-Dev-MTP-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

17,73914
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/diffusiongemma-26B-A4B-it-GGUF

unsloth/diffusiongemma-26B-A4B-it-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

17,577394
26.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

llm-jp/llm-jp-4-33b-thinking

llm-jp/llm-jp-4-33b-thinking is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

17,54339
33.0B80 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Sohailhosseini/Qwen3.8-9B-Distill-FP8

Sohailhosseini/Qwen3.8-9B-Distill-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

17,4840
9.0B24 GB+ VRAMapache-2.0
Deployment details
sentence-similarity

antoinelouis/colbert-xm

antoinelouis/colbert-xm is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

17,42971
Unknown4 GB+ VRAMmit
Deployment details
text-generation

z-lab/Qwen3.5-9B-DFlash

z-lab/Qwen3.5-9B-DFlash is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

17,31143
9.0B24 GB+ VRAMapache-2.0
Deployment details
depth-estimation

google/tipsv2-b14-dpt

google/tipsv2-b14-dpt is a depth estimation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

17,01514
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

mlx-community/Qwen3.5-9B-OptiQ-4bit

mlx-community/Qwen3.5-9B-OptiQ-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

16,93896
9.0B8 GB+ VRAMapache-2.0
Deployment details
text-generation

ibm-granite/granite-4.2-3b

ibm-granite/granite-4.2-3b is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

16,72176
3.0B12 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

lightonai/LightOnOCR-2-1B-bbox-soup

lightonai/LightOnOCR-2-1B-bbox-soup is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

16,69518
1.0B6 GB+ VRAMapache-2.0
Deployment details