HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

192,823 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

fill-mask

google-bert/bert-base-cased

google-bert/bert-base-cased is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,640,600371
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

lmstudio-community/Qwen3.8-27B-MLX-5bit

lmstudio-community/Qwen3.8-27B-MLX-5bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

3,639,5200
27.0B64 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen2.5-7B-Instruct-AWQ

Qwen/Qwen2.5-7B-Instruct-AWQ is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

3,593,56551
7.0B24 GB+ VRAMapache-2.0
Deployment details
text-classification

BAAI/bge-reranker-base

BAAI/bge-reranker-base is a text classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,580,098242
Unknown8 GB+ VRAMmit
Deployment details
general AI

Qwen/Qwen3-TTS-12Hz-1.7B-Base

Qwen/Qwen3-TTS-12Hz-1.7B-Base is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

3,558,093507
1.7B6 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen3-4B-Instruct-2507

Qwen/Qwen3-4B-Instruct-2507 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

3,538,174955
4.0B12 GB+ VRAMapache-2.0
Deployment details
text-generation

EleutherAI/pythia-160m

EleutherAI/pythia-160m is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,443,53445
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen3-1.7B

Qwen/Qwen3-1.7B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

3,441,144531
1.7B6 GB+ VRAMapache-2.0
Deployment details
text-generation

ornith-ai/Ornith-1.5-35B-A3B-GGUF

ornith-ai/Ornith-1.5-35B-A3B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

3,430,123388
35.0B32 GB+ VRAMmit
Deployment details
automatic-speech-recognition

jonatasgrosman/wav2vec2-large-xlsr-53-greek

jonatasgrosman/wav2vec2-large-xlsr-53-greek is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,294,6594
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

unsloth/Qwen3.8-27B-NVFP4

unsloth/Qwen3.8-27B-NVFP4 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

3,263,102434
27.0B64 GB+ VRAMapache-2.0
Deployment details
zero-shot-classification

facebook/bart-large-mnli

facebook/bart-large-mnli is a zero shot classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,251,4651,606
Unknown8 GB+ VRAMmit
Deployment details
any-to-any

google/gemma-4-E2B-it

google/gemma-4-E2B-it is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,222,423946
Unknown8 GB+ VRAMapache-2.0
Deployment details
voice-activity-detectionGated

pyannote/segmentation

pyannote/segmentation is a voice activity detection model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.

3,198,851694
Unknown8 GB+ VRAMmit
Deployment details
automatic-speech-recognition

Qwen/Qwen3-ASR-1.7B

Qwen/Qwen3-ASR-1.7B is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

3,161,8321,076
1.7B6 GB+ VRAMapache-2.0
Deployment details
general AI

charactr/vocos-mel-24khz

charactr/vocos-mel-24khz is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,146,06943
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

Qwen/Qwen3.5-2B

Qwen/Qwen3.5-2B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,115,741384
2.0B8 GB+ VRAMapache-2.0
Deployment details
any-to-any

google/gemma-4-12B-it

google/gemma-4-12B-it is a any to any model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

3,087,4031,531
12.0B32 GB+ VRAMapache-2.0
Deployment details
image-feature-extraction

facebook/dinov2-base

facebook/dinov2-base is a image feature extraction model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,081,152195
Unknown8 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

openai/whisper-small

openai/whisper-small is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,035,223595
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Qwen/Qwen3-VL-2B-Instruct

Qwen/Qwen3-VL-2B-Instruct is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,986,461460
2.0B8 GB+ VRAMapache-2.0
Deployment details
fill-mask

microsoft/deberta-v3-base

microsoft/deberta-v3-base is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,948,155440
Unknown8 GB+ VRAMmit
Deployment details
general AI

facebook/wav2vec2-base

facebook/wav2vec2-base is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,919,894125
Unknown8 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognitionGated

pyannote/voice-activity-detection

pyannote/voice-activity-detection is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.

2,866,839241
Unknown8 GB+ VRAMmit
Deployment details