HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

198,823 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

fill-mask

emilyalsentzer/Bio_ClinicalBERT

emilyalsentzer/Bio_ClinicalBERT is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,193,352440
Unknown8 GB+ VRAMmit
Deployment details
text-generation

Qwen/Qwen2.5-32B-Instruct

Qwen/Qwen2.5-32B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

2,181,370358
32.0B80 GB+ VRAMapache-2.0
Deployment details
feature-extraction

Qwen/Qwen3-Embedding-8B

Qwen/Qwen3-Embedding-8B is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,166,281798
8.0B24 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

Systran/faster-whisper-tiny

Systran/faster-whisper-tiny is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,153,80126
Unknown8 GB+ VRAMmit
Deployment details
zero-shot-image-classification

patrickjohncyh/fashion-clip

patrickjohncyh/fashion-clip is a zero shot image classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,146,617291
Unknown8 GB+ VRAMmit
Deployment details
automatic-speech-recognition

jonatasgrosman/wav2vec2-large-xlsr-53-persian

jonatasgrosman/wav2vec2-large-xlsr-53-persian is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,146,05429
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

RadixArk/Qwen3.8-27B-NVFP4

RadixArk/Qwen3.8-27B-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

2,144,88688
27.0B64 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen2.5-14B-Instruct-AWQ

Qwen/Qwen2.5-14B-Instruct-AWQ is a text generation model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.

2,136,38737
14.0B48 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

Khalsuu/filipino-wav2vec2-l-xls-r-300m-official

Khalsuu/filipino-wav2vec2-l-xls-r-300m-official is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,130,8002
Unknown8 GB+ VRAMapache-2.0
Deployment details
time-series-forecasting

autogluon/chronos-bolt-base

autogluon/chronos-bolt-base is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,108,78034
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

speechbrain/spkrec-ecapa-voxceleb

speechbrain/spkrec-ecapa-voxceleb is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,104,029344
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen3-30B-A3B

Qwen/Qwen3-30B-A3B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

2,097,577939
30.0B80 GB+ VRAMapache-2.0
Deployment details
text-generation

antirez/deepseek-v4-gguf

antirez/deepseek-v4-gguf is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,088,647468
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

Qwen/Qwen2-VL-2B-Instruct

Qwen/Qwen2-VL-2B-Instruct is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,037,704518
2.0B8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

zai-org/GLM-OCR

zai-org/GLM-OCR is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,024,7922,019
Unknown8 GB+ VRAMmit
Deployment details
feature-extraction

BAAI/bge-base-en

BAAI/bge-base-en is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

1,998,80762
Unknown4 GB+ VRAMmit
Deployment details
feature-extraction

intfloat/multilingual-e5-large-instruct

intfloat/multilingual-e5-large-instruct is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

1,980,898635
Unknown4 GB+ VRAMmit
Deployment details
text-generation

Qwen/Qwen3-4B-Base

Qwen/Qwen3-4B-Base is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

1,967,14598
4.0B12 GB+ VRAMapache-2.0
Deployment details
general AI

facebook/esmfold_v1

facebook/esmfold_v1 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,966,84753
Unknown8 GB+ VRAMmit
Deployment details
audio-to-audio

nvidia/bigvgan_v2_22khz_80band_256x

nvidia/bigvgan_v2_22khz_80band_256x is a audio to audio model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,957,21633
Unknown8 GB+ VRAMmit
Deployment details
token-classification

dslim/bert-base-NER

dslim/bert-base-NER is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,940,352732
Unknown8 GB+ VRAMmit
Deployment details