HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

1,676,984 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

apodex/Apodex-1.1-mini-GPTQ-Int4

apodex/Apodex-1.1-mini-GPTQ-Int4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,151♡ 18
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-generation

svc-nai-cci/nanollama-public

svc-nai-cci/nanollama-public is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,150♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →
general AI

mradermacher/Zora-9B-v2-i1-GGUF

mradermacher/Zora-9B-v2-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,150♡ 1
9.0B8 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

Vontra/Qwen3.8-27B-MLX-8bit

Vontra/Qwen3.8-27B-MLX-8bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

↓ 1,150♡ 1
27.0B32 GB+ VRAMapache-2.0
Deployment details →
text-generation

unsloth/Qwen3-0.6B-FP8

unsloth/Qwen3-0.6B-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 1,149♡ 1
600M4 GB+ VRAMapache-2.0
Deployment details →
general AI

Atotti/Qwen3-Omni-AudioTransformer

Atotti/Qwen3-Omni-AudioTransformer is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,149♡ 37
Unknown8 GB+ VRAMapache-2.0
Deployment details →
general AI

Mungert/MAI-UI-8B-GGUF

Mungert/MAI-UI-8B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,149♡ 0
8.0B8 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

llamaindex/vdr-2b-multi-v1

llamaindex/vdr-2b-multi-v1 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,149♡ 128
2.0B8 GB+ VRAMapache-2.0
Deployment details →