HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

204,822 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

zai-org/GLM-5.2

zai-org/GLM-5.2 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

965,2545,077
Unknown8 GB+ VRAMmit
Deployment details
image-classification

timm/resnet50.ram_in1k

timm/resnet50.ram_in1k is a image classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

964,4520
Unknown8 GB+ VRAMapache-2.0
Deployment details
token-classification

LocalAI-io/privacy-filter-nemotron-GGUF

LocalAI-io/privacy-filter-nemotron-GGUF is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

963,0070
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

deepseek-ai/DeepSeek-OCR-2

deepseek-ai/DeepSeek-OCR-2 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

962,8611,088
Unknown8 GB+ VRAMapache-2.0
Deployment details
feature-extraction

michaelfeil/bge-small-en-v1.5

michaelfeil/bge-small-en-v1.5 is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

962,5033
Unknown4 GB+ VRAMmit
Deployment details
general AI

circlestone-labs/Anima

circlestone-labs/Anima is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

961,2792,197
Unknown8 GB+ VRAMother
Deployment details
sentence-similarity

Alibaba-NLP/gte-large-en-v1.5

Alibaba-NLP/gte-large-en-v1.5 is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

960,309239
Unknown4 GB+ VRAMapache-2.0
Deployment details
feature-extraction

microsoft/wavlm-base-plus

microsoft/wavlm-base-plus is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

956,10640
Unknown4 GB+ VRAMLicense unknown
Deployment details
image-text-to-text

RedHatAI/Qwen3.6-35B-A3B-NVFP4

RedHatAI/Qwen3.6-35B-A3B-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

953,068173
35.0B80 GB+ VRAMapache-2.0
Deployment details
text-generation

ornith-ai/Ornith-1.5-397B-GGUF

ornith-ai/Ornith-1.5-397B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.

951,99935
397.0B256 GB+ VRAMmit
Deployment details
automatic-speech-recognition

nvidia/nemotron-3.5-asr-streaming-0.6b

nvidia/nemotron-3.5-asr-streaming-0.6b is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

949,7191,091
600M4 GB+ VRAMother
Deployment details
feature-extraction

laion/clap-htsat-unfused

laion/clap-htsat-unfused is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

947,75879
Unknown4 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

handy-computer/cohere-transcribe-03-2026-gguf

handy-computer/cohere-transcribe-03-2026-gguf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

945,3753
Unknown8 GB+ VRAMapache-2.0
Deployment details
object-detection

PaddlePaddle/PP-DocLayoutV3_safetensors

PaddlePaddle/PP-DocLayoutV3_safetensors is a object detection model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

944,75439
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

facebook/wav2vec2-xls-r-300m

facebook/wav2vec2-xls-r-300m is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

940,148134
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen3-0.6B-Base

Qwen/Qwen3-0.6B-Base is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

939,807189
600M4 GB+ VRAMapache-2.0
Deployment details
text-generation

nvidia/GLM-5.2-NVFP4

nvidia/GLM-5.2-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

939,693321
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

unsloth/Qwen3.8-Flash-Next-GGUF

unsloth/Qwen3.8-Flash-Next-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

935,568836
Unknown8 GB+ VRAMother
Deployment details
image-text-to-text

unsloth/Qwen3.5-4B-GGUF

unsloth/Qwen3.5-4B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

928,147408
4.0B6 GB+ VRAMapache-2.0
Deployment details
feature-extraction

jinaai/jina-embeddings-v2-small-en

jinaai/jina-embeddings-v2-small-en is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

927,637142
Unknown4 GB+ VRAMapache-2.0
Deployment details