HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

1,023,045 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

aj9o9/Qwen3.6-35B-A3B-Escha-W2-GGUF

aj9o9/Qwen3.6-35B-A3B-Escha-W2-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

↓ 1,508♡ 9
35.0B32 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

FadedRedStar/Qwen3.5-4B-heretic-GGUF

FadedRedStar/Qwen3.5-4B-heretic-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

↓ 1,507♡ 2
4.0B6 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

pauleb/Qwen3.6-27B-Q4_K_M-GGUF

pauleb/Qwen3.6-27B-Q4_K_M-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 1,506♡ 0
27.0B24 GB+ VRAMapache-2.0
Deployment details →
general AI

mradermacher/SP-7B-i1-GGUF

mradermacher/SP-7B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,505♡ 0
7.0B8 GB+ VRAMmit
Deployment details →
text-generation

mlx-community/MiniCPM5-1B-OptiQ-4bit

mlx-community/MiniCPM5-1B-OptiQ-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 1,505♡ 10
1.0B4 GB+ VRAMapache-2.0
Deployment details →
fill-mask

latincy/latin-bert

latincy/latin-bert is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,505♡ 1
Unknown8 GB+ VRAMapache-2.0
Deployment details →
any-to-any

FastFlowLM/Gemma4-E4B-IT-NPU2

FastFlowLM/Gemma4-E4B-IT-NPU2 is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,504♡ 2
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-to-speech

cstr/dots-tts-soar-GGUF

cstr/dots-tts-soar-GGUF is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,502♡ 3
Unknown8 GB+ VRAMapache-2.0
Deployment details →
general AI

mradermacher/Goetia-24B-v1.4-GGUF

mradermacher/Goetia-24B-v1.4-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 1,502♡ 1
24.0B24 GB+ VRAMapache-2.0
Deployment details →