HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

937,050 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

aj9o9/GLM-5.3-Flash-GGUF

aj9o9/GLM-5.3-Flash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,781♡ 8
Unknown8 GB+ VRAMmit
Deployment details →
text-ranking

NeuML/biomedbert-base-reranker

NeuML/biomedbert-base-reranker is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,777♡ 6
Unknown8 GB+ VRAMapache-2.0
Deployment details →
general AI

mradermacher/mox-8b-i1-GGUF

mradermacher/mox-8b-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,777♡ 0
8.0B8 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

surogate/Qwen3.5-0.8B-FP8

surogate/Qwen3.5-0.8B-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 1,776♡ 1
800M4 GB+ VRAMapache-2.0
Deployment details →
general AI

mradermacher/NaNovel-9B-i1-GGUF

mradermacher/NaNovel-9B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,774♡ 2
9.0B8 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

avlp12/Qwen3.8-27B-Alis-MLX-4bit

avlp12/Qwen3.8-27B-Alis-MLX-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 1,774♡ 4
27.0B24 GB+ VRAMapache-2.0
Deployment details →
general AI

kernels-community/paged-attention

kernels-community/paged-attention is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,774♡ 11
Unknown8 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

unsloth/gemma-4-E4B-it-MLX-8bit

unsloth/gemma-4-E4B-it-MLX-8bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,773♡ 12
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-to-image

meituan-longcat/LongCat-Image-Dev

meituan-longcat/LongCat-Image-Dev is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

↓ 1,772♡ 50
Unknown12 GB+ VRAMapache-2.0
Deployment details →
text-generation

TaoLiveAIGC/TLive-Omni-4B

TaoLiveAIGC/TLive-Omni-4B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

↓ 1,772♡ 17
4.0B12 GB+ VRAMapache-2.0
Deployment details →
general AI

mradermacher/Asita-8B-i1-GGUF

mradermacher/Asita-8B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,771♡ 1
8.0B8 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

vs4vijay/Muse-Glimmer-30B-GGUF

vs4vijay/Muse-Glimmer-30B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 1,771♡ 0
30.0B24 GB+ VRAMapache-2.0
Deployment details →
text-generation

CodeDevX/MultiModel-Small-229M

CodeDevX/MultiModel-Small-229M is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,771♡ 0
Unknown8 GB+ VRAMmit
Deployment details →
image-text-to-text

aoiandroid/gemma-4-E2B-it-GGUF

aoiandroid/gemma-4-E2B-it-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,771♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →
fill-mask

ctheodoris/Geneformer

ctheodoris/Geneformer is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,770♡ 311
Unknown8 GB+ VRAMapache-2.0
Deployment details →