HUGGING FACE MODEL INDEXSearch and compare Hugging Face models.
795,858 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
general AI
mradermacher/Ministral-3-14B-Reasoning-2512-SOM-MPOA-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
↓ 2,161♡ 1
14.0B12 GB+ VRAMapache-2.0
Deployment details →
general AI
mradermacher/Qwen3.8-2B-Heretic-Max-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
↓ 2,161♡ 1
2.0B4 GB+ VRAMapache-2.0
Deployment details →
image-to-image
tonera/FLUX.2-klein-4B-fp8-diffusers is a image to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
↓ 2,160♡ 1
4.0B12 GB+ VRAMapache-2.0
Deployment details →
general AI
JMingo/gemma-4-12B-it-heretic-UD-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
↓ 2,160♡ 10
12.0B12 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text
Sohailhosseini/Qwen3.8-9B-Distill-AWQ-W4A16 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 2,160♡ 0
9.0B24 GB+ VRAMapache-2.0
Deployment details →
text-generation
FreedomAISVR/GLM-4.7-Flash-NVFP4-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 2,159♡ 2
Unknown8 GB+ VRAMmit
Deployment details →
general AI
snuh/hari-q3-8b is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 2,158♡ 6
8.0B24 GB+ VRAMapache-2.0
Deployment details →
general AI
dahara1/gemma-4-E4B-it-UD-japanese-imatrix is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 2,158♡ 1
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-generation
ibm-granite/granite-4.1-3b-fp8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
↓ 2,157♡ 8
3.0B4 GB+ VRAMapache-2.0
Deployment details →
text-generation
ibm-granite/granite-8b-code-base-4k is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 2,157♡ 32
8.0B24 GB+ VRAMapache-2.0
Deployment details →
text-generation
bartowski/allenai_SERA-8B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 2,156♡ 0
8.0B8 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text
kingjones777/Ornith-1.5-35B-A3B-ROCmFPX-AGENT-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
↓ 2,156♡ 4
35.0B32 GB+ VRAMapache-2.0
Deployment details →
image-text-to-image
drbaph/HiDream-O1-Image-Dev-2604-BF16 is a image text to image model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 2,155♡ 2
Unknown8 GB+ VRAMmit
Deployment details →
text-generation
bartowski/janhq_Jan-code-4b-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
↓ 2,154♡ 2
4.0B6 GB+ VRAMapache-2.0
Deployment details →
general AI
mradermacher/Hypernova-60B-2602-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
↓ 2,153♡ 2
60.0B48 GB+ VRAMapache-2.0
Deployment details →
text-generation
puwaer/DeepSeek-V4-Flash-0731-reap-200b-gguf is a text generation model indexed for deployment research. Estimated minimum GPU memory is 192 GB. It is publicly listed on Hugging Face.
↓ 2,153♡ 3
200.0B192 GB+ VRAMmit
Deployment details →
image-text-to-text
software-mansion/react-native-executorch-qwen-3.5 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 2,150♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text
loolzrulez/gemma-4-26B-A4B-Q4_K_M-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 2,149♡ 0
26.0B24 GB+ VRAMapache-2.0
Deployment details →
general AI
mradermacher/Muse-Glimmer-30B-heretic-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 2,149♡ 2
30.0B24 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text
pipenetwork/GLM-5.3-Flash-MLX-6bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 2,149♡ 3
Unknown8 GB+ VRAMmit
Deployment details →
audio-to-audio
aufklarer/Sidon-CoreML is a audio to audio model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 2,148♡ 0
Unknown8 GB+ VRAMmit
Deployment details →
text-generation
igorls/gemma-4-12B-it-qat-q4_0-unquantized-heretic-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
↓ 2,147♡ 16
12.0B12 GB+ VRAMapache-2.0
Deployment details →
feature-extraction
Anbeeld/Qwen3.5-4B-DFlash-GGUF is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
↓ 2,147♡ 3
4.0B6 GB+ VRAMapache-2.0
Deployment details →
general AI
mradermacher/GLM-4.7-Flash-Derestricted-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 2,146♡ 11
Unknown8 GB+ VRAMmit
Deployment details →