HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

462,288 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

DreamFast/Qwen3-VL-8B-Heretic-1.3.0

DreamFast/Qwen3-VL-8B-Heretic-1.3.0 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,13714
8.0B8 GB+ VRAMapache-2.0
Deployment details
text-generation

drawais/Qwen3-Embedding-4B-AWQ-INT4

drawais/Qwen3-Embedding-4B-AWQ-INT4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

3,1335
4.0B6 GB+ VRAMapache-2.0
Deployment details
any-to-any

unsloth/gemma-4-E4B-it-qat-w4a16

unsloth/gemma-4-E4B-it-qat-w4a16 is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,1313
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

Jamphus/PinkCherry_NSFW_LTX23

Jamphus/PinkCherry_NSFW_LTX23 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,1222
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

CrossNow/Qwen3.8-27B-GGUF

CrossNow/Qwen3.8-27B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

3,1210
27.0B24 GB+ VRAMapache-2.0
Deployment details
general AI

amd/gpt-oss-20b-WFP8-AFP8-KVFP8

amd/gpt-oss-20b-WFP8-AFP8-KVFP8 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.

3,1200
20.0B48 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

FastFlowLM/Qwen3.6-35B-A3B-NPU2

FastFlowLM/Qwen3.6-35B-A3B-NPU2 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

3,12015
35.0B80 GB+ VRAMapache-2.0
Deployment details
text-generation

ai9stars/G9v3-3B

ai9stars/G9v3-3B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

3,11958
3.0B12 GB+ VRAMapache-2.0
Deployment details
text-generation

mradermacher/sarv-reasoning-i1-GGUF

mradermacher/sarv-reasoning-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,1090
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

turboderp/Qwen3.8-27B-exl3

turboderp/Qwen3.8-27B-exl3 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

3,10895
27.0B64 GB+ VRAMapache-2.0
Deployment details
text-generation

ThorOdinson246/nl2sh-1.5b-Q4_K_M

ThorOdinson246/nl2sh-1.5b-Q4_K_M is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

3,08858
1.5B4 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

zeromodels/qwen3-vl-2b-thinking

zeromodels/qwen3-vl-2b-thinking is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,0800
2.0B8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

True2456/Qwen3.8-27B-AWQ-4.85bpw

True2456/Qwen3.8-27B-AWQ-4.85bpw is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

3,07211
27.0B64 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

OpenVINO/gemma-4-E2B-it-int4-ov

OpenVINO/gemma-4-E2B-it-int4-ov is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,0666
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

furiosa-ai/gpt-oss-20b

furiosa-ai/gpt-oss-20b is a text generation model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.

3,0600
20.0B48 GB+ VRAMapache-2.0
Deployment details