HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

188,825 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

image-text-to-text

Qwen/Qwen3.8-27B-FP8

Qwen/Qwen3.8-27B-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

6,985,465770
27.0B64 GB+ VRAMapache-2.0
Deployment details
text-generation

farbodtavakkoli/OTel-2.0-LLM-31B-IT

farbodtavakkoli/OTel-2.0-LLM-31B-IT is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

6,970,01315
31.0B80 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

openai/whisper-large-v3-turbo

openai/whisper-large-v3-turbo is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,902,9953,304
Unknown8 GB+ VRAMmit
Deployment details
sentence-similarity

intfloat/multilingual-e5-base

intfloat/multilingual-e5-base is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

6,897,165383
Unknown4 GB+ VRAMmit
Deployment details
image-text-to-text

Qwen/Qwen3.8-27B

Qwen/Qwen3.8-27B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

6,712,16014,355
27.0B64 GB+ VRAMapache-2.0
Deployment details
text-generation

openai/gpt-oss-20b

openai/gpt-oss-20b is a text generation model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.

6,551,1915,007
20.0B48 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen2.5-0.5B-Instruct

Qwen/Qwen2.5-0.5B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

6,428,381620
500M4 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen3-4B

Qwen/Qwen3-4B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

6,283,681694
4.0B12 GB+ VRAMapache-2.0
Deployment details
time-series-forecasting

autogluon/chronos-bolt-small

autogluon/chronos-bolt-small is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,145,86961
Unknown8 GB+ VRAMapache-2.0
Deployment details
fill-mask

FacebookAI/roberta-large

FacebookAI/roberta-large is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,121,295319
Unknown8 GB+ VRAMmit
Deployment details
time-series-forecasting

autogluon/chronos-2-small

autogluon/chronos-2-small is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,957,7756
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-ranking

cross-encoder/ms-marco-MiniLM-L4-v2

cross-encoder/ms-marco-MiniLM-L4-v2 is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,875,40229
Unknown8 GB+ VRAMapache-2.0
Deployment details
voice-activity-detectionGated

pyannote/segmentation-3.0

pyannote/segmentation-3.0 is a voice activity detection model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.

5,859,4741,704
Unknown8 GB+ VRAMmit
Deployment details
image-classification

google/vit-base-patch16-224

google/vit-base-patch16-224 is a image classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,480,705996
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

openai/gpt-oss-120b

openai/gpt-oss-120b is a text generation model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.

5,414,1625,172
120.0B256 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen3-32B

Qwen/Qwen3-32B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

5,156,288743
32.0B80 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

openai/whisper-large-v3

openai/whisper-large-v3 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,106,1406,252
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-to-video

Comfy-Org/Wan_2.2_ComfyUI_Repackaged

Comfy-Org/Wan_2.2_ComfyUI_Repackaged is a image to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,060,055864
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Qwen/Qwen3.6-27B

Qwen/Qwen3.6-27B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

4,876,2602,287
27.0B64 GB+ VRAMapache-2.0
Deployment details
feature-extraction

BAAI/bge-small-zh-v1.5

BAAI/bge-small-zh-v1.5 is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

4,862,103139
Unknown4 GB+ VRAMmit
Deployment details
any-to-any

google/gemma-4-E4B-it

google/gemma-4-E4B-it is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,805,3951,536
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

dphn/dolphin-2.9.1-yi-1.5-34b

dphn/dolphin-2.9.1-yi-1.5-34b is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

4,769,15265
34.0B80 GB+ VRAMapache-2.0
Deployment details