HUGGING FACE MODEL INDEXSearch and compare Hugging Face models.
480,281 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
image-text-to-text
nightmedia/Qwen3.8-27B-Cold-Fusion-GAIN-V1.1-Fable-mxfp8-mlx is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
↓ 2,831♡ 3
27.0B64 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text
mobilint/Qwen3-VL-8B-Instruct-Batch16 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 2,827♡ 0
8.0B24 GB+ VRAMapache-2.0
Deployment details →
sentence-similarity
VAGOsolutions/SauerkrautLM-Reason-EuroColBERT is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
↓ 2,827♡ 12
Unknown4 GB+ VRAMapache-2.0
Deployment details →
general AI
offmonreal/Qwen3.8-27B-MaxQuality-iMatrix-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 2,825♡ 3
27.0B24 GB+ VRAMapache-2.0
Deployment details →
text-generation
TrevorJS/gemma-4-E4B-it-uncensored is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 2,822♡ 37
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-generation
hotdogs/Ornith-1.0-9B-abliterated-fable-MTP-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 2,819♡ 0
9.0B8 GB+ VRAMmit
Deployment details →
text-generation
saidutta69/MiniCPM5-1B-heretic is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
↓ 2,817♡ 3
1.0B4 GB+ VRAMapache-2.0
Deployment details →
text-generation
ibm-granite/granite-20b-code-instruct-8k is a text generation model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
↓ 2,816♡ 44
20.0B48 GB+ VRAMapache-2.0
Deployment details →
text-to-image
SeeSee21/Z-Anime is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
↓ 2,816♡ 485
Unknown12 GB+ VRAMapache-2.0
Deployment details →
general AI
mradermacher/Qwen3.5-4B-abliterated-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
↓ 2,815♡ 2
4.0B6 GB+ VRAMapache-2.0
Deployment details →
general AI
Kraekin/Goetia-26B-A4B-v1.3-Absolute-Heretic-ARA-Q4_K_S-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 2,813♡ 1
26.0B24 GB+ VRAMapache-2.0
Deployment details →
general AI
SC117/gemma-4-E4B-it-heretic-QAT-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 2,812♡ 20
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-generation
ibm-granite/granite-4.1-30b-fp8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 2,811♡ 7
30.0B24 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text
bartowski/InternScience_Agents-A1-4B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
↓ 2,810♡ 15
4.0B6 GB+ VRAMapache-2.0
Deployment details →
text-to-speech
cstr/kokoro-voices-GGUF is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 2,808♡ 2
Unknown8 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text
bartowski/kai-os_Carnice-V3-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 2,808♡ 2
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-generation
CPSPX/babylm-zho-pinyin-code-97M is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 2,807♡ 0
Unknown8 GB+ VRAMmit
Deployment details →
general AI
mradermacher/Asmodeus-24B-v3-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 2,806♡ 3
24.0B24 GB+ VRAMapache-2.0
Deployment details →
general AI
mradermacher/Qwen3-VL-Embedding-2B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
↓ 2,805♡ 2
2.0B4 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text
Atomic-Germ/Ornith-1.0-9B-NPU2 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 2,805♡ 0
9.0B8 GB+ VRAMmit
Deployment details →
image-text-to-text
Hatim2221/Mubsir-Qwen-2B-VL is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 2,801♡ 4
2.0B8 GB+ VRAMapache-2.0
Deployment details →
token-classification
gliner-community/gliner_small-v2.5 is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 2,800♡ 14
Unknown8 GB+ VRAMapache-2.0
Deployment details →
general AI
Janvitos/gemma-4-12B-it-qat-assistant-MTP-Q8_0-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
↓ 2,800♡ 38
12.0B12 GB+ VRAMapache-2.0
Deployment details →
text-generation
mlx-community/KAT-Coder-V2.5-Dev-OptiQ-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 2,800♡ 17
Unknown8 GB+ VRAMapache-2.0
Deployment details →