HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

202,823 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

sentence-similarity

Qdrant/all-MiniLM-L6-v2-onnx

Qdrant/all-MiniLM-L6-v2-onnx is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

1,155,8237
Unknown4 GB+ VRAMapache-2.0
Deployment details
image-feature-extraction

timm/vit_small_patch14_dinov2.lvd142m

timm/vit_small_patch14_dinov2.lvd142m is a image feature extraction model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,138,9248
Unknown8 GB+ VRAMapache-2.0
Deployment details
feature-extraction

BAAI/bge-large-zh-v1.5

BAAI/bge-large-zh-v1.5 is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

1,128,186646
Unknown4 GB+ VRAMmit
Deployment details
sentence-similarity

Qwen/Qwen3-VL-Embedding-8B

Qwen/Qwen3-VL-Embedding-8B is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

1,109,399476
8.0B24 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

Systran/faster-whisper-large-v3

Systran/faster-whisper-large-v3 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,105,706655
Unknown8 GB+ VRAMmit
Deployment details
image-classification

microsoft/resnet-50

microsoft/resnet-50 is a image classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,105,417508
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen3-4B-Instruct-2507-FP8

Qwen/Qwen3-4B-Instruct-2507-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

1,100,66082
4.0B12 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

QuantTrio/Qwen3.5-9B-AWQ

QuantTrio/Qwen3.5-9B-AWQ is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

1,086,59427
9.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

raxcore-dev/Rax-4.5

raxcore-dev/Rax-4.5 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,085,8655
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

allenai/OLMo-2-0425-1B

allenai/OLMo-2-0425-1B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

1,085,68881
1.0B6 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

Yehor/w2v-xls-r-uk

Yehor/w2v-xls-r-uk is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,082,5538
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/Muse-Glimmer-30B-GGUF

unsloth/Muse-Glimmer-30B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

1,075,366522
30.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

openai-community/gpt2-large

openai-community/gpt2-large is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,073,821358
Unknown8 GB+ VRAMmit
Deployment details
sentence-similarity

intfloat/e5-base-v2

intfloat/e5-base-v2 is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

1,072,365157
Unknown4 GB+ VRAMmit
Deployment details
text-to-speech

Qwen/Qwen3-TTS-12Hz-0.6B-CustomVoice

Qwen/Qwen3-TTS-12Hz-0.6B-CustomVoice is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

1,065,889183
600M4 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen2.5-1.5B

Qwen/Qwen2.5-1.5B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

1,063,114215
1.5B6 GB+ VRAMapache-2.0
Deployment details
sentence-similarity

thenlper/gte-small

thenlper/gte-small is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

1,061,404190
Unknown4 GB+ VRAMmit
Deployment details
text-generation

OBLITERATUS/Qwen3.8-27B-OBLITERATED

OBLITERATUS/Qwen3.8-27B-OBLITERATED is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

1,059,3871,135
27.0B24 GB+ VRAMapache-2.0
Deployment details