HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

200,823 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

nvidia/Gemma-4-26B-A4B-NVFP4

nvidia/Gemma-4-26B-A4B-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

1,742,109136
26.0B64 GB+ VRAMapache-2.0
Deployment details
general AI

docling-project/docling-layout-heron

docling-project/docling-layout-heron is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,720,56853
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

nvidia/Gemma-4-31B-IT-NVFP4

nvidia/Gemma-4-31B-IT-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

1,718,321563
31.0B80 GB+ VRAMother
Deployment details
automatic-speech-recognition

Harveenchadha/vakyansh-wav2vec2-tamil-tam-250

Harveenchadha/vakyansh-wav2vec2-tamil-tam-250 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,715,8114
Unknown8 GB+ VRAMmit
Deployment details
automatic-speech-recognition

openai/whisper-base

openai/whisper-base is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,712,755287
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/Qwen3.6-35B-A3B-NVFP4

unsloth/Qwen3.6-35B-A3B-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

1,709,688117
35.0B80 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen3-14B

Qwen/Qwen3-14B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.

1,703,411463
14.0B48 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen3-Coder-Next-FP8

Qwen/Qwen3-Coder-Next-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,701,361179
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-classification

timm/resnet18.a1_in1k

timm/resnet18.a1_in1k is a image classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,699,47614
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen2.5-0.5B

Qwen/Qwen2.5-0.5B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

1,691,482443
500M4 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

airesearch/wav2vec2-large-xlsr-53-th

airesearch/wav2vec2-large-xlsr-53-th is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,681,50628
Unknown8 GB+ VRAMcc-by-sa-4.0
Deployment details
image-text-to-text

Qwen/Qwen2.5-VL-32B-Instruct-AWQ

Qwen/Qwen2.5-VL-32B-Instruct-AWQ is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

1,679,91365
32.0B80 GB+ VRAMapache-2.0
Deployment details
text-ranking

cross-encoder/ms-marco-MiniLM-L12-v2

cross-encoder/ms-marco-MiniLM-L12-v2 is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,679,599111
Unknown8 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

Systran/faster-whisper-base

Systran/faster-whisper-base is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,677,27435
Unknown8 GB+ VRAMmit
Deployment details
zero-shot-image-classification

openai/clip-vit-base-patch16

openai/clip-vit-base-patch16 is a zero shot image classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,661,896166
Unknown8 GB+ VRAMLicense unknown
Deployment details
zero-shot-object-detection

IDEA-Research/grounding-dino-base

IDEA-Research/grounding-dino-base is a zero shot object detection model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,652,763206
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-textGated

google/gemma-3-4b-it

google/gemma-3-4b-it is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. Access approval is required on Hugging Face.

1,645,9911,476
4.0B12 GB+ VRAMgemma
Deployment details
automatic-speech-recognition

mesolitica/wav2vec2-xls-r-300m-mixed

mesolitica/wav2vec2-xls-r-300m-mixed is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,645,6435
Unknown8 GB+ VRAMLicense unknown
Deployment details
image-to-videoGated

Lightricks/LTX-2.5

Lightricks/LTX-2.5 is a image to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.

1,644,7963,172
Unknown8 GB+ VRAMother
Deployment details
image-text-to-text

Qwen/Qwen2-VL-7B-Instruct-AWQ

Qwen/Qwen2-VL-7B-Instruct-AWQ is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

1,634,14948
7.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

TinyLlama/TinyLlama-1.1B-Chat-v1.0

TinyLlama/TinyLlama-1.1B-Chat-v1.0 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

1,626,5671,776
1.1B6 GB+ VRAMapache-2.0
Deployment details