HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

216,821 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

empero-ai/Qwen3.8-9B-Distill

empero-ai/Qwen3.8-9B-Distill is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

25,355195
9.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

cyankiwi/Qwen3.5-2B-AWQ-4bit

cyankiwi/Qwen3.5-2B-AWQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

25,1983
2.0B4 GB+ VRAMapache-2.0
Deployment details
text-generation

swiss-ai/Apertus-70B-Instruct-2509

swiss-ai/Apertus-70B-Instruct-2509 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 192 GB. It is publicly listed on Hugging Face.

25,042194
70.0B192 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/gemma-4-31B-it

unsloth/gemma-4-31B-it is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

24,98321
31.0B80 GB+ VRAMapache-2.0
Deployment details
feature-extraction

codefuse-ai/F2LLM-v2-80M

codefuse-ai/F2LLM-v2-80M is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

24,46114
Unknown4 GB+ VRAMapache-2.0
Deployment details
any-to-any

google/gemma-4-12B-it-assistant

google/gemma-4-12B-it-assistant is a any to any model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

24,434121
12.0B32 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

RedHatAI/Qwen3.8-27B-NVFP4

RedHatAI/Qwen3.8-27B-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

24,39711
27.0B64 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

protoLabsAI/Ornith-1.5-9B-MTP-GGUF

protoLabsAI/Ornith-1.5-9B-MTP-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

24,28025
9.0B8 GB+ VRAMmit
Deployment details
text-generation

FINAL-Bench/Ourbox-35B-JGOS-GGUF

FINAL-Bench/Ourbox-35B-JGOS-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

24,12230
35.0B32 GB+ VRAMapache-2.0
Deployment details
text-generation

inclusionAI/Ling-3.0-tiny

inclusionAI/Ling-3.0-tiny is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

23,836406
Unknown8 GB+ VRAMmit
Deployment details
text-generation

6block/Qwen3-32B-GGUF

6block/Qwen3-32B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

23,8001
32.0B24 GB+ VRAMapache-2.0
Deployment details
general AI

Comfy-Org/lotus

Comfy-Org/lotus is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

23,6805
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

Comfy-Org/void-model

Comfy-Org/void-model is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

23,57126
Unknown8 GB+ VRAMapache-2.0
Deployment details