HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

1,153,035 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

RedHatAI/Qwen3-4B-FP8-dynamic

RedHatAI/Qwen3-4B-FP8-dynamic is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

↓ 1,403♡ 1
4.0B12 GB+ VRAMapache-2.0
Deployment details →
general AI

Barrrrry/DeepSeek-R1-W4AFP8

Barrrrry/DeepSeek-R1-W4AFP8 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,402♡ 3
Unknown8 GB+ VRAMmit
Deployment details →
image-text-to-text

lyssquant/Inkling-Small-GGUF

lyssquant/Inkling-Small-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,402♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-generation

mradermacher/OpenGCM-v2-i1-GGUF

mradermacher/OpenGCM-v2-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,402♡ 1
Unknown8 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

Vontra/Qwen3.8-27B-oQ2

Vontra/Qwen3.8-27B-oQ2 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

↓ 1,402♡ 2
27.0B64 GB+ VRAMapache-2.0
Deployment details →
sentence-similarity

BAAI/BGE-VL-large

BAAI/BGE-VL-large is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 1,400♡ 25
Unknown4 GB+ VRAMmit
Deployment details →
text-generation

evsinlb/Qwen3.8-27B-oQ4e-mtp

evsinlb/Qwen3.8-27B-oQ4e-mtp is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 1,400♡ 1
27.0B24 GB+ VRAMapache-2.0
Deployment details →
token-classification

piotrmaciejbednarski/gliner2-polish-pii

piotrmaciejbednarski/gliner2-polish-pii is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,399♡ 1
Unknown8 GB+ VRAMapache-2.0
Deployment details →
automatic-speech-recognition

cstr/wav2vec2-large-xlsr-53-english-GGUF

cstr/wav2vec2-large-xlsr-53-english-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,398♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-generation

mradermacher/ChatBerry-i1-GGUF

mradermacher/ChatBerry-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,397♡ 1
Unknown8 GB+ VRAMapache-2.0
Deployment details →