HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

368,302 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

RedHatAI/GLM-5.2-NVFP4-FP8

RedHatAI/GLM-5.2-NVFP4-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,50824
Unknown8 GB+ VRAMmit
Deployment details
text-generation

litert-community/Qwen3-4B

litert-community/Qwen3-4B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

6,4909
4.0B12 GB+ VRAMapache-2.0
Deployment details
mask-generation

tiiuae/Falcon-Perception

tiiuae/Falcon-Perception is a mask generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,481143
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

lightonai/LightOnOCR-2-1B-base

lightonai/LightOnOCR-2-1B-base is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

6,46814
1.0B6 GB+ VRAMapache-2.0
Deployment details
feature-extraction

Anbeeld/Qwen3.6-35B-A3B-DFlash-GGUF

Anbeeld/Qwen3.6-35B-A3B-DFlash-GGUF is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

6,46412
35.0B32 GB+ VRAMapache-2.0
Deployment details
general AI

mradermacher/Decka-4B-i1-GGUF

mradermacher/Decka-4B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

6,4600
4.0B6 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

mlx-community/Qwen3.8-27B-nvfp4

mlx-community/Qwen3.8-27B-nvfp4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

6,4508
27.0B64 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

PassingByPixels/Qwen3.8-27B-NVFP4

PassingByPixels/Qwen3.8-27B-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

6,4462
27.0B64 GB+ VRAMapache-2.0
Deployment details
text-generation

wangzhang/GLM-4.7-Flash-abliteratex

wangzhang/GLM-4.7-Flash-abliteratex is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,4323
Unknown8 GB+ VRAMmit
Deployment details
reinforcement-learning

mradermacher/KernelBench-RLVR-120b-i1-GGUF

mradermacher/KernelBench-RLVR-120b-i1-GGUF is a reinforcement learning model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

6,4230
120.0B80 GB+ VRAMapache-2.0
Deployment details
text-generation

rodrigoramosrs/veriloop-coder-e1-gguf

rodrigoramosrs/veriloop-coder-e1-gguf is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,42114
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-ranking

onebrain-ai/onebrain-rerank-v1

onebrain-ai/onebrain-rerank-v1 is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,4151
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

mlx-community/Qwen3.8-27B-mxfp8

mlx-community/Qwen3.8-27B-mxfp8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

6,4097
27.0B64 GB+ VRAMapache-2.0
Deployment details
text-generation

wangzhang/Qwen3.8-27B-abliterated

wangzhang/Qwen3.8-27B-abliterated is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

6,39411
27.0B64 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

thoaibuiic/PhoWhisper-large-ct2

thoaibuiic/PhoWhisper-large-ct2 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,3930
Unknown8 GB+ VRAMmit
Deployment details
general AI

fcreait/Qwen-Image-Edit-mflux

fcreait/Qwen-Image-Edit-mflux is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,3831
Unknown8 GB+ VRAMapache-2.0
Deployment details
zero-shot-image-classification

google/tipsv2-l14

google/tipsv2-l14 is a zero shot image classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,36923
Unknown8 GB+ VRAMapache-2.0
Deployment details