HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

380,302 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

unsloth/GLM-5.3-Flash-FP8

unsloth/GLM-5.3-Flash-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,73824
Unknown8 GB+ VRAMmit
Deployment details
text-generation

abenzerps/Spark-X2.5-4B-GGUF

abenzerps/Spark-X2.5-4B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

5,7029
4.0B6 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognitionGated

ARTPARK-IISc/SraVaani-1.0

ARTPARK-IISc/SraVaani-1.0 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.

5,68176
Unknown8 GB+ VRAMmit
Deployment details
text-to-image

byteshape/Qwen-Image-2512-GGUF

byteshape/Qwen-Image-2512-GGUF is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

5,6777
Unknown12 GB+ VRAMapache-2.0
Deployment details
general AI

kandinskylab/KVAE-3D-2.0-t4s16

kandinskylab/KVAE-3D-2.0-t4s16 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,66812
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

z-lab/Qwen3.5-35B-A3B-DFlash

z-lab/Qwen3.5-35B-A3B-DFlash is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

5,66545
35.0B80 GB+ VRAMapache-2.0
Deployment details
general AI

mudler/Nex-N2-mini-APEX-GGUF

mudler/Nex-N2-mini-APEX-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,64711
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-classification

fastino/gliguard-LLMGuardrails-300M

fastino/gliguard-LLMGuardrails-300M is a text classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,64182
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

ibm-granite/granitelib-core-r1.0

ibm-granite/granitelib-core-r1.0 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,61831
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

prism-ml/Ternary-Bonsai-4B-gguf

prism-ml/Ternary-Bonsai-4B-gguf is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

5,60734
4.0B6 GB+ VRAMapache-2.0
Deployment details
text-generation

DreamFast/qwen3-4b-heretic

DreamFast/qwen3-4b-heretic is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

5,59335
4.0B6 GB+ VRAMapache-2.0
Deployment details