HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

198,823 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

Qwen/Qwen3-14B-AWQ

Qwen/Qwen3-14B-AWQ is a text generation model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.

2,303,43173
14.0B48 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Qwen/Qwen3.5-27B

Qwen/Qwen3.5-27B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

2,296,7921,040
27.0B64 GB+ VRAMapache-2.0
Deployment details
sentence-similarity

nomic-ai/nomic-embed-text-v2-moe

nomic-ai/nomic-embed-text-v2-moe is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

2,293,748498
Unknown4 GB+ VRAMapache-2.0
Deployment details
feature-extraction

jinaai/jina-embeddings-v3

jinaai/jina-embeddings-v3 is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

2,290,8501,154
Unknown4 GB+ VRAMcc-by-nc-4.0
Deployment details
image-text-to-text

Qwen/Qwen3.5-35B-A3B

Qwen/Qwen3.5-35B-A3B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

2,290,2361,500
35.0B80 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

gigant/romanian-wav2vec2

gigant/romanian-wav2vec2 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,279,1378
Unknown8 GB+ VRAMapache-2.0
Deployment details
feature-extraction

facebook/w2v-bert-2.0

facebook/w2v-bert-2.0 is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

2,271,473228
Unknown4 GB+ VRAMmit
Deployment details
general AI

docling-project/docling-models

docling-project/docling-models is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,269,046216
Unknown8 GB+ VRAMcdla-permissive-2.0
Deployment details
image-text-to-text

cyankiwi/gemma-4-26B-A4B-it-AWQ-4bit

cyankiwi/gemma-4-26B-A4B-it-AWQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,266,35296
26.0B24 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

anuragshas/wav2vec2-large-xlsr-53-telugu

anuragshas/wav2vec2-large-xlsr-53-telugu is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,241,7375
Unknown8 GB+ VRAMapache-2.0
Deployment details
sentence-similarityGated

google/embeddinggemma-300m

google/embeddinggemma-300m is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. Access approval is required on Hugging Face.

2,225,7171,891
Unknown4 GB+ VRAMgemma
Deployment details
automatic-speech-recognition

KBLab/wav2vec2-large-voxrex-swedish

KBLab/wav2vec2-large-voxrex-swedish is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,221,14413
Unknown8 GB+ VRAMcc0-1.0
Deployment details
fill-mask

emilyalsentzer/Bio_ClinicalBERT

emilyalsentzer/Bio_ClinicalBERT is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,193,352440
Unknown8 GB+ VRAMmit
Deployment details
text-generation

Qwen/Qwen2.5-32B-Instruct

Qwen/Qwen2.5-32B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

2,181,370358
32.0B80 GB+ VRAMapache-2.0
Deployment details
feature-extraction

Qwen/Qwen3-Embedding-8B

Qwen/Qwen3-Embedding-8B is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,166,281798
8.0B24 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

Systran/faster-whisper-tiny

Systran/faster-whisper-tiny is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,153,80126
Unknown8 GB+ VRAMmit
Deployment details
zero-shot-image-classification

patrickjohncyh/fashion-clip

patrickjohncyh/fashion-clip is a zero shot image classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,146,617291
Unknown8 GB+ VRAMmit
Deployment details
automatic-speech-recognition

jonatasgrosman/wav2vec2-large-xlsr-53-persian

jonatasgrosman/wav2vec2-large-xlsr-53-persian is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,146,05429
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

RadixArk/Qwen3.8-27B-NVFP4

RadixArk/Qwen3.8-27B-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

2,144,88688
27.0B64 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen2.5-14B-Instruct-AWQ

Qwen/Qwen2.5-14B-Instruct-AWQ is a text generation model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.

2,136,38737
14.0B48 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

Khalsuu/filipino-wav2vec2-l-xls-r-300m-official

Khalsuu/filipino-wav2vec2-l-xls-r-300m-official is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,130,8002
Unknown8 GB+ VRAMapache-2.0
Deployment details
time-series-forecasting

autogluon/chronos-bolt-base

autogluon/chronos-bolt-base is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,108,78034
Unknown8 GB+ VRAMapache-2.0
Deployment details