HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

442,292 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

litert-community/Qwen3-0.6B-int4

litert-community/Qwen3-0.6B-int4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

3,4901
600M4 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

IAAR-Shanghai/Metis-9B

IAAR-Shanghai/Metis-9B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

3,4775
9.0B24 GB+ VRAMapache-2.0
Deployment details
general AI

Tele-AI/TeleChat2-3B

Tele-AI/TeleChat2-3B is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

3,4763
3.0B12 GB+ VRAMapache-2.0
Deployment details
translation

NiuTrans/LMT-60-0.6B

NiuTrans/LMT-60-0.6B is a translation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

3,47210
600M4 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

pipenetwork/GLM-5.3-Flash-MLX-8bit

pipenetwork/GLM-5.3-Flash-MLX-8bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,4724
Unknown8 GB+ VRAMmit
Deployment details
general AI

mradermacher/AFM-4.5B-i1-GGUF

mradermacher/AFM-4.5B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

3,4701
4.5B6 GB+ VRAMapache-2.0
Deployment details
text-generation

ibm-granite/granite-4.1-8b-base

ibm-granite/granite-4.1-8b-base is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

3,46626
8.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

shivanandasai/Qwen3.8-27B-OBLITERATED

shivanandasai/Qwen3.8-27B-OBLITERATED is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

3,4640
27.0B24 GB+ VRAMapache-2.0
Deployment details
general AI

mradermacher/Reelva-12B-i1-GGUF

mradermacher/Reelva-12B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

3,4640
12.0B12 GB+ VRAMapache-2.0
Deployment details
text-generation

lukealonso/GLM-5.2-NVFP4

lukealonso/GLM-5.2-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,46132
Unknown8 GB+ VRAMmit
Deployment details