HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

200,823 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

nvidia/Gemma-4-26B-A4B-NVFP4

nvidia/Gemma-4-26B-A4B-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

1,742,109136
26.0B64 GB+ VRAMapache-2.0
Deployment details
general AI

docling-project/docling-layout-heron

docling-project/docling-layout-heron is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,720,56853
Unknown8 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

Harveenchadha/vakyansh-wav2vec2-tamil-tam-250

Harveenchadha/vakyansh-wav2vec2-tamil-tam-250 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,715,8114
Unknown8 GB+ VRAMmit
Deployment details
automatic-speech-recognition

openai/whisper-base

openai/whisper-base is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,712,755287
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/Qwen3.6-35B-A3B-NVFP4

unsloth/Qwen3.6-35B-A3B-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

1,709,688117
35.0B80 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen3-14B

Qwen/Qwen3-14B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.

1,703,411463
14.0B48 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen3-Coder-Next-FP8

Qwen/Qwen3-Coder-Next-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,701,361179
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-classification

timm/resnet18.a1_in1k

timm/resnet18.a1_in1k is a image classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,699,47614
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen2.5-0.5B

Qwen/Qwen2.5-0.5B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

1,691,482443
500M4 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Qwen/Qwen2.5-VL-32B-Instruct-AWQ

Qwen/Qwen2.5-VL-32B-Instruct-AWQ is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

1,679,91365
32.0B80 GB+ VRAMapache-2.0
Deployment details
text-ranking

cross-encoder/ms-marco-MiniLM-L12-v2

cross-encoder/ms-marco-MiniLM-L12-v2 is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,679,599111
Unknown8 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

Systran/faster-whisper-base

Systran/faster-whisper-base is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,677,27435
Unknown8 GB+ VRAMmit
Deployment details
zero-shot-object-detection

IDEA-Research/grounding-dino-base

IDEA-Research/grounding-dino-base is a zero shot object detection model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,652,763206
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Qwen/Qwen2-VL-7B-Instruct-AWQ

Qwen/Qwen2-VL-7B-Instruct-AWQ is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

1,634,14948
7.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

TinyLlama/TinyLlama-1.1B-Chat-v1.0

TinyLlama/TinyLlama-1.1B-Chat-v1.0 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

1,626,5671,776
1.1B6 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

openai/whisper-tiny

openai/whisper-tiny is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,624,155443
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/Qwen3.5-9B-GGUF

unsloth/Qwen3.5-9B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,618,479889
9.0B8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

vikhyatk/moondream2

vikhyatk/moondream2 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,615,1841,435
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Qwen/Qwen3.5-122B-A10B-FP8

Qwen/Qwen3.5-122B-A10B-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.

1,612,865115
122.0B256 GB+ VRAMapache-2.0
Deployment details
general AI

google/mobilebert-uncased

google/mobilebert-uncased is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,610,04175
Unknown8 GB+ VRAMapache-2.0
Deployment details