HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

605,071 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

general AI

mradermacher/Holo-3.1-4B-GGUF

mradermacher/Holo-3.1-4B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

5,1747
4.0B6 GB+ VRAMapache-2.0
Deployment details
general AI

lapizkxra/nxappdd

lapizkxra/nxappdd is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,1731
Unknown8 GB+ VRAMLicense unknown
Deployment details
general AI

bullerwins/Kimi-K3-GGUF

bullerwins/Kimi-K3-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,1710
Unknown8 GB+ VRAMLicense unknown
Deployment details
text-generation

mradermacher/Qwen3.8-Queen-27B-GGUF

mradermacher/Qwen3.8-Queen-27B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

5,1692
27.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

RedHatAI/GLM-5.2-MXFP4xFP8_BLOCK

RedHatAI/GLM-5.2-MXFP4xFP8_BLOCK is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,1660
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

stepfun-ai/Step-3.7-Flash-GGUF

stepfun-ai/Step-3.7-Flash-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,159172
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

drawais/Qwen3-Reranker-0.6B-AWQ-INT4

drawais/Qwen3-Reranker-0.6B-AWQ-INT4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

5,1580
600M4 GB+ VRAMapache-2.0
Deployment details
text-generation

reaperdoesntknow/TameForCasualLM

reaperdoesntknow/TameForCasualLM is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,1550
Unknown8 GB+ VRAMapache-2.0
Deployment details
sentence-similarity

LiquidAI/LFM2.5-Embedding-350M-GGUF

LiquidAI/LFM2.5-Embedding-350M-GGUF is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

5,15238
Unknown4 GB+ VRAMother
Deployment details
image-text-to-text

unsloth/Qwen3.5-0.8B-Base

unsloth/Qwen3.5-0.8B-Base is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

5,1403
800M4 GB+ VRAMapache-2.0
Deployment details
text-generation

z-lab/Muse-Glimmer-30B-DFlash2-GGUF

z-lab/Muse-Glimmer-30B-DFlash2-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

5,13712
30.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

rostlabs/rost-1b-instruct

rostlabs/rost-1b-instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

5,1321
1.0B6 GB+ VRAMcc-by-nc-4.0
Deployment details
image-text-to-text

QQZ2026/Qwen3.8-27B-NVFP4-Q5K-no-MTP-GGUF

QQZ2026/Qwen3.8-27B-NVFP4-Q5K-no-MTP-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

5,1266
27.0B24 GB+ VRAMLicense unknown
Deployment details