HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

366,302 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

z-lab/Qwen3.5-4B-DFlash

z-lab/Qwen3.5-4B-DFlash is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

6,74940
4.0B12 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

cstr/cohere-transcribe-03-2026-GGUF

cstr/cohere-transcribe-03-2026-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,7379
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

amd/Qwen3.5-397B-A17B-MXFP4-AttnFP8-V2

amd/Qwen3.5-397B-A17B-MXFP4-AttnFP8-V2 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.

6,7280
397.0B256 GB+ VRAMapache-2.0
Deployment details
text-to-speech

SPRINGLab/Indic-Mio

SPRINGLab/Indic-Mio is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,64628
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-textGated

orcarouter/GLM-5.3-Flash-Uncensored-GGUF

orcarouter/GLM-5.3-Flash-Uncensored-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.

6,61030
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

numind/NuExtract-2.0-8B

numind/NuExtract-2.0-8B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

6,59464
8.0B24 GB+ VRAMmit
Deployment details
video-classification

apiantonio/vjepa2.1-vit-base-384

apiantonio/vjepa2.1-vit-base-384 is a video classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,5742
Unknown8 GB+ VRAMmit
Deployment details
text-generation

logic65/Qwen3.8-Whittle-16B

logic65/Qwen3.8-Whittle-16B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 16 GB. It is publicly listed on Hugging Face.

6,5552
16.0B16 GB+ VRAMapache-2.0
Deployment details
text-generation

ordlibrary/hauhau-qwen36-uncensored

ordlibrary/hauhau-qwen36-uncensored is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,5200
Unknown8 GB+ VRAMmit
Deployment details
text-generation

RedHatAI/GLM-5.2-NVFP4-FP8

RedHatAI/GLM-5.2-NVFP4-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,50824
Unknown8 GB+ VRAMmit
Deployment details
text-generation

litert-community/Qwen3-4B

litert-community/Qwen3-4B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

6,4909
4.0B12 GB+ VRAMapache-2.0
Deployment details
mask-generation

tiiuae/Falcon-Perception

tiiuae/Falcon-Perception is a mask generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,481143
Unknown8 GB+ VRAMapache-2.0
Deployment details