HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

661,069 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-to-video

TaoLiveAIGC/TaoMate

TaoLiveAIGC/TaoMate is a text to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,3213
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-to-speech

cstr/chatterbox-turbo-GGUF

cstr/chatterbox-turbo-GGUF is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,3192
Unknown8 GB+ VRAMmit
Deployment details
text-generation

marin-dna/marin-dna-exp135-m5.1

marin-dna/marin-dna-exp135-m5.1 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,3192
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

ibm-granite/granite-vision-3.2-2b

ibm-granite/granite-vision-3.2-2b is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,316125
2.0B8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/gemma-4-26b-a4b-it-UD-MLX-4bit

unsloth/gemma-4-26b-a4b-it-UD-MLX-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,31534
26.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

JetBrains/Mellum2-12B-A2.5B-Thinking

JetBrains/Mellum2-12B-A2.5B-Thinking is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

2,311335
12.0B32 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

dots-studio/dots3-note-prev

dots-studio/dots3-note-prev is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,311269
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

WhiskyAKM/Ling-3.0-Tiny-GGUF

WhiskyAKM/Ling-3.0-Tiny-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,3103
Unknown8 GB+ VRAMmit
Deployment details
sentence-similarity

NeuML/bert-tiny-sts-last-pooling

NeuML/bert-tiny-sts-last-pooling is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

2,3091
Unknown4 GB+ VRAMapache-2.0
Deployment details
general AI

mradermacher/A2R-30B-A3B-i1-GGUF

mradermacher/A2R-30B-A3B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,3080
30.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

reaperdoesntknow/SMOLM2Prover-GGUF

reaperdoesntknow/SMOLM2Prover-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,3080
Unknown8 GB+ VRAMapache-2.0
Deployment details
visual-document-retrieval

tencent/EVIE-Preview-4.5B

tencent/EVIE-Preview-4.5B is a visual document retrieval model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

2,30796
4.5B12 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

bartowski/Jackrong_Qwen3.5-9B-Neo-GGUF

bartowski/Jackrong_Qwen3.5-9B-Neo-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,3068
9.0B8 GB+ VRAMapache-2.0
Deployment details
text-generation

sallani/ISO27001-Qwen2.5-0.5B-Edge

sallani/ISO27001-Qwen2.5-0.5B-Edge is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

2,3060
500M4 GB+ VRAMapache-2.0
Deployment details
general AI

Elena2810/z-image-turbo-q4_k_s

Elena2810/z-image-turbo-q4_k_s is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,3050
Unknown8 GB+ VRAMapache-2.0
Deployment details