HUGGING FACE MODEL INDEXSearch and compare Hugging Face models.
1,051,042 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
general AI
mradermacher/Total04-DeepSeek-R1-Distill-Llama-70B-heretic-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
↓ 1,475♡ 3
70.0B48 GB+ VRAMmit
Deployment details →
image-text-to-text
mlx-community/gemma-4-12B-coder-fable5-composer2.5-v1-OptiQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
↓ 1,475♡ 8
12.0B12 GB+ VRAMapache-2.0
Deployment details →
text-generation
mradermacher/MiniCPM5-1B-Base-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
↓ 1,475♡ 0
1.0B4 GB+ VRAMapache-2.0
Deployment details →
image-to-text
mradermacher/Lh41-1042-Magellanic-7B-0711-i1-GGUF is a image to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 1,474♡ 0
7.0B8 GB+ VRAMapache-2.0
Deployment details →
image-to-text
mradermacher/Hulu-Med-30A3-i1-GGUF is a image to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 1,474♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →
general AI
mradermacher/Gliese-Qwen3.5-27B-Abliterated-Caption-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 1,474♡ 0
27.0B24 GB+ VRAMapache-2.0
Deployment details →
general AI
mradermacher/ERNIE-4.5-HUGE-51B-A3B-Thinking-Brainstorm40x-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
↓ 1,473♡ 0
51.0B48 GB+ VRAMapache-2.0
Deployment details →
general AI
mradermacher/OpenAI-gpt-oss-20B-INSTRUCT-Heretic-Uncensored-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 16 GB. It is publicly listed on Hugging Face.
↓ 1,473♡ 0
20.0B16 GB+ VRAMapache-2.0
Deployment details →
text-generation
RedHatAI/Mistral-Small-24B-Instruct-2501-FP8-dynamic is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
↓ 1,473♡ 13
24.0B64 GB+ VRAMapache-2.0
Deployment details →
text-generation
bartowski/XiaomiMiMo_MiMo-V2-Flash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 1,472♡ 10
Unknown8 GB+ VRAMmit
Deployment details →
general AI
mradermacher/Impish_Bloodmoon_12B_Abliterated-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
↓ 1,472♡ 3
12.0B12 GB+ VRAMapache-2.0
Deployment details →
text-generation
longtermrisk/Llama-3.1-8B-target-only-no-hallucination-first-third-sft-seed2-epoch3 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 1,472♡ 0
8.0B24 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text
nightmedia/Qwen3.8-27B-Holodeck-mxfp4-mlx is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
↓ 1,472♡ 2
27.0B64 GB+ VRAMapache-2.0
Deployment details →
general AI
mradermacher/GPT-OSS-26B-abliterated-Preview-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 1,471♡ 1
26.0B24 GB+ VRAMapache-2.0
Deployment details →
general AI
mradermacher/IoGPT-A1-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 1,471♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-generation
Xkev/gemma-3-1b-it-kk is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
↓ 1,471♡ 1
1.0B6 GB+ VRAMmit
Deployment details →
feature-extraction
biohub/ESMC-6B-sae-layer60-k64-codebook16384 is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 16 GB. It is publicly listed on Hugging Face.
↓ 1,471♡ 1
6.0B16 GB+ VRAMmit
Deployment details →
text-generation
mradermacher/Ministral-3-14B-Reasoning-2512-Uncensored-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
↓ 1,471♡ 1
14.0B12 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text
shisa-ai/Qwen3.8-27B-FP8-BLOCK is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
↓ 1,471♡ 1
27.0B64 GB+ VRAMapache-2.0
Deployment details →
general AI
mradermacher/Ostrich-27B-Qwen3.8-260816-Abliterated-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 1,471♡ 0
27.0B24 GB+ VRAMapache-2.0
Deployment details →
text-generation
AtomicChat/gemma-4-26B-A4B-it-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 1,470♡ 1
26.0B24 GB+ VRAMapache-2.0
Deployment details →
text-to-speech
Audio8/Audio8-TTS-Preview-0.6B-ONNX-INT4 is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
↓ 1,470♡ 54
600M4 GB+ VRAMapache-2.0
Deployment details →
text-generation
hotdogs/Qwen3.8-27B-thinkingcap-abliterated is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
↓ 1,470♡ 2
27.0B64 GB+ VRAMapache-2.0
Deployment details →
general AI
BasedAGI/Qwen2.5-32B-AGI-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 1,469♡ 1
32.0B24 GB+ VRAMmit
Deployment details →