HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

623,071 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

ai-sage/GigaChat3.1-10B-A1.8B-GGUF

ai-sage/GigaChat3.1-10B-A1.8B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

4,81286
10.0B12 GB+ VRAMmit
Deployment details
image-text-to-text

unsloth/gemma-4-E4B-it-UD-MLX-4bit

unsloth/gemma-4-E4B-it-UD-MLX-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,80745
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-to-image

AlperKTS/Krea-2-SVDQuant-ComfyUI

AlperKTS/Krea-2-SVDQuant-ComfyUI is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

4,80217
Unknown12 GB+ VRAMother
Deployment details
text-to-speech

herimor/voxtream2

herimor/voxtream2 is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,7978
Unknown8 GB+ VRAMcc-by-4.0
Deployment details
image-text-to-text

furiosa-ai/Qwen3-VL-4B-Thinking

furiosa-ai/Qwen3-VL-4B-Thinking is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

4,7970
4.0B12 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Vishva007/Qwen3.8-27B-W4A16-AutoRound

Vishva007/Qwen3.8-27B-W4A16-AutoRound is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

4,7862
27.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

nvidia/MiniMax-M3-DSpark

nvidia/MiniMax-M3-DSpark is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,78413
Unknown8 GB+ VRAMother
Deployment details
text-generation

furiosa-ai/Qwen3-8B-FP8

furiosa-ai/Qwen3-8B-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

4,7840
8.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

byteshape/Qwen3.6-35B-A3B-GGUF

byteshape/Qwen3.6-35B-A3B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

4,77939
35.0B32 GB+ VRAMapache-2.0
Deployment details
text-generation

poolside/Laguna-XS-2.1-FP8

poolside/Laguna-XS-2.1-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,7799
Unknown8 GB+ VRAMopenmdw-1.1
Deployment details
automatic-speech-recognition

handy-computer/whisper-tiny-gguf

handy-computer/whisper-tiny-gguf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,7741
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

mPLUG/GUI-Owl-1.5-8B-Instruct

mPLUG/GUI-Owl-1.5-8B-Instruct is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

4,77113
8.0B24 GB+ VRAMmit
Deployment details