HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

1,065,041 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

FreedomAISVR/Qwen3.5-9B-NVFP4-GGUF

FreedomAISVR/Qwen3.5-9B-NVFP4-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,442♡ 0
9.0B8 GB+ VRAMapache-2.0
Deployment details →
text-generation

mradermacher/next2.5-i1-GGUF

mradermacher/next2.5-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,441♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-generation

mrs83/Kurtis-EON1-Hybrid-2B-v0.1.2

mrs83/Kurtis-EON1-Hybrid-2B-v0.1.2 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,440♡ 0
2.0B8 GB+ VRAMapache-2.0
Deployment details →
general AI

mradermacher/squeez-2b-i1-GGUF

mradermacher/squeez-2b-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 1,439♡ 2
2.0B4 GB+ VRAMapache-2.0
Deployment details →
fill-mask

zymonody/chinese-babylm-v4

zymonody/chinese-babylm-v4 is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,439♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →
image-to-image

w3ss/Qwen-Image-Edit-2511-GGUF

w3ss/Qwen-Image-Edit-2511-GGUF is a image to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

↓ 1,439♡ 1
Unknown12 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

QuaduxIT/Qwen3.6-35B-A3B-QD-GGUF

QuaduxIT/Qwen3.6-35B-A3B-QD-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

↓ 1,439♡ 0
35.0B32 GB+ VRAMapache-2.0
Deployment details →
general AI

mradermacher/Serenity-27B-i1-GGUF

mradermacher/Serenity-27B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 1,439♡ 0
27.0B24 GB+ VRAMapache-2.0
Deployment details →
automatic-speech-recognition

nyadla-sys/whisper-tiny.en.tflite

nyadla-sys/whisper-tiny.en.tflite is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,438♡ 1
Unknown8 GB+ VRAMmit
Deployment details →
video-text-to-text

yanziang/InternVideo3-8B-Instruct

yanziang/InternVideo3-8B-Instruct is a video text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 1,438♡ 9
8.0B24 GB+ VRAMapache-2.0
Deployment details →
summarization

mradermacher/turbo-ai-7b-i1-GGUF

mradermacher/turbo-ai-7b-i1-GGUF is a summarization model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,438♡ 2
7.0B8 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

armandosds/gemma-4-E4B-it-GGUF

armandosds/gemma-4-E4B-it-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,438♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →