HUGGING FACE MODEL INDEXSearch and compare Hugging Face models.
432,300 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
text-generation
kernelogic/Qwen3.8-27B-Uncensored-GPTQ-Int4-sym-G128-MTP-BF16 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 3,637♡ 1
27.0B24 GB+ VRAMapache-2.0
Deployment details →
general AI
mradermacher/gemma-4-12B-agentic-fable5-composer2.5-v2-3.5x-tau2-uncensored-heretic-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
↓ 3,636♡ 11
12.0B12 GB+ VRAMapache-2.0
Deployment details →
general AI
protoLabsAI/ThinkingCap-Qwen3.6-27B-MTP-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 3,635♡ 65
27.0B24 GB+ VRAMapache-2.0
Deployment details →
text-generation
lmcoleman/Qwen3.6-27B-Fable-Fusion-711-MTP-ROCmFPX-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 3,635♡ 5
27.0B24 GB+ VRAMapache-2.0
Deployment details →
general AI
mradermacher/Katarau-9B-ru-RP-nsfw-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 3,632♡ 7
9.0B8 GB+ VRAMapache-2.0
Deployment details →
text-generation
RedHatAI/TinyLlama-1.1B-Chat-v1.0-pruned2.4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
↓ 3,625♡ 2
1.1B6 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text
llmfan46/gemma-4-31B-it-qat-q4_0-uncensored-heretic-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 3,620♡ 36
31.0B24 GB+ VRAMapache-2.0
Deployment details →
general AI
Zyphra/ZAYA1-74B-preview-legacy is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 192 GB. It is publicly listed on Hugging Face.
↓ 3,620♡ 0
74.0B192 GB+ VRAMapache-2.0
Deployment details →
general AI
deucebucket/Qwen3-30B-A3B-Cerebellum-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 3,614♡ 2
30.0B24 GB+ VRAMapache-2.0
Deployment details →
any-to-any
unsloth/gemma-4-12B-it-qat-q4_0-unquantized is a any to any model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
↓ 3,614♡ 9
12.0B12 GB+ VRAMapache-2.0
Deployment details →
general AI
mradermacher/gemma-4-E2B-it-Uncensored-MAX-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 3,611♡ 6
Unknown8 GB+ VRAMapache-2.0
Deployment details →
zero-shot-image-classification
srpone/zooclaw-fashionsiglip2 is a zero shot image classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 3,603♡ 17
Unknown8 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text
cloudbjorn/Qwen3.8-27B-Yes-Man-uncensored is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
↓ 3,598♡ 1
27.0B64 GB+ VRAMapache-2.0
Deployment details →
token-classification
Kansallisarkisto/finbert-ner is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 3,597♡ 3
Unknown8 GB+ VRAMmit
Deployment details →
text-generation
nazeshinjite/DeepSeek-V4-Flash-0731-ds4-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 3,596♡ 21
Unknown8 GB+ VRAMmit
Deployment details →
text-generation
Jackrong/MLX-Qwen3.5-9B-DeepSeek-V4-Flash-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 3,595♡ 17
9.0B8 GB+ VRAMapache-2.0
Deployment details →
text-generation
DavidAU/Gemma-3-it-4B-Uncensored-DBL-X-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
↓ 3,594♡ 68
4.0B6 GB+ VRAMapache-2.0
Deployment details →
text-generation
AMAImedia/GLM-5.3-Flash is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 3,591♡ 0
Unknown8 GB+ VRAMmit
Deployment details →
text-generation
cyankiwi/granite-4.0-h-micro-AWQ-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 3,588♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →
sentence-similarity
mixedbread-ai/mxbai-edge-colbert-v0-17m is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
↓ 3,588♡ 38
Unknown4 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text
kkuspa/Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-MTP-NVFP4A16 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
↓ 3,576♡ 8
27.0B64 GB+ VRAMapache-2.0
Deployment details →
text-generation
lmstudio-community/Ornith-1.0-35B-MLX-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
↓ 3,573♡ 2
35.0B32 GB+ VRAMmit
Deployment details →
text-generation
zaakirio/Ornith-1.5-9B-Uncensored-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 3,573♡ 5
9.0B8 GB+ VRAMmit
Deployment details →
general AI
mradermacher/Ornith-1.5-9B-heretic-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 3,572♡ 1
9.0B8 GB+ VRAMmit
Deployment details →