HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

663,069 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

reaperdoesntknow/SMOLM2Prover-GGUF

reaperdoesntknow/SMOLM2Prover-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,3080
Unknown8 GB+ VRAMapache-2.0
Deployment details
visual-document-retrieval

tencent/EVIE-Preview-4.5B

tencent/EVIE-Preview-4.5B is a visual document retrieval model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

2,30796
4.5B12 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

bartowski/Jackrong_Qwen3.5-9B-Neo-GGUF

bartowski/Jackrong_Qwen3.5-9B-Neo-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,3068
9.0B8 GB+ VRAMapache-2.0
Deployment details
text-generation

sallani/ISO27001-Qwen2.5-0.5B-Edge

sallani/ISO27001-Qwen2.5-0.5B-Edge is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

2,3060
500M4 GB+ VRAMapache-2.0
Deployment details
general AI

Elena2810/z-image-turbo-q4_k_s

Elena2810/z-image-turbo-q4_k_s is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,3050
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

Atomic-Germ/Gemma4-E4B-GLM-NPU2

Atomic-Germ/Gemma4-E4B-GLM-NPU2 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,3050
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

giladgd/gemma-4-31B-it-GGUF

giladgd/gemma-4-31B-it-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,2990
31.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Chungulus/Qwen3.8-27B-IQ4_NL-GGUF

Chungulus/Qwen3.8-27B-IQ4_NL-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,2993
27.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

litert-community/Qwen3-8B

litert-community/Qwen3-8B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,29810
8.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

ggml-org/MiniCPM-V-4.6-GGUF

ggml-org/MiniCPM-V-4.6-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,29628
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

litert-community/VibeThinker-1.5B

litert-community/VibeThinker-1.5B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

2,2954
1.5B6 GB+ VRAMmit
Deployment details
image-text-to-text

Chungulus/Qwen3.8-27B-Q2_K-GGUF

Chungulus/Qwen3.8-27B-Q2_K-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,2942
27.0B24 GB+ VRAMapache-2.0
Deployment details