HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

188,825 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

Qwen/Qwen2.5-1.5B-Instruct

Qwen/Qwen2.5-1.5B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

7,326,700820
1.5B6 GB+ VRAMapache-2.0
Deployment details
zero-shot-image-classification

openai/clip-vit-large-patch14

openai/clip-vit-large-patch14 is a zero shot image classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

7,308,3432,074
Unknown8 GB+ VRAMLicense unknown
Deployment details
general AI

Comfy-Org/z_image_turbo

Comfy-Org/z_image_turbo is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

7,289,306870
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Qwen/Qwen3.5-4B

Qwen/Qwen3.5-4B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

7,264,578897
4.0B12 GB+ VRAMapache-2.0
Deployment details
text-to-speech

coqui/XTTS-v2

coqui/XTTS-v2 is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

7,241,4313,772
Unknown8 GB+ VRAMother
Deployment details
image-text-to-text

Qwen/Qwen3.6-27B-FP8

Qwen/Qwen3.6-27B-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

7,198,015354
27.0B64 GB+ VRAMapache-2.0
Deployment details
fill-mask

distilbert/distilbert-base-uncased

distilbert/distilbert-base-uncased is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

7,138,1521,170
Unknown8 GB+ VRAMapache-2.0
Deployment details
feature-extraction

intfloat/multilingual-e5-large

intfloat/multilingual-e5-large is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

6,999,9251,248
Unknown4 GB+ VRAMmit
Deployment details
image-text-to-text

Qwen/Qwen3.8-27B-FP8

Qwen/Qwen3.8-27B-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

6,985,465770
27.0B64 GB+ VRAMapache-2.0
Deployment details
text-generation

farbodtavakkoli/OTel-2.0-LLM-31B-IT

farbodtavakkoli/OTel-2.0-LLM-31B-IT is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

6,970,01315
31.0B80 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

openai/whisper-large-v3-turbo

openai/whisper-large-v3-turbo is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,902,9953,304
Unknown8 GB+ VRAMmit
Deployment details
sentence-similarity

intfloat/multilingual-e5-base

intfloat/multilingual-e5-base is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

6,897,165383
Unknown4 GB+ VRAMmit
Deployment details
image-text-to-text

Qwen/Qwen3.8-27B

Qwen/Qwen3.8-27B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

6,712,16014,355
27.0B64 GB+ VRAMapache-2.0
Deployment details
text-generation

openai/gpt-oss-20b

openai/gpt-oss-20b is a text generation model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.

6,551,1915,007
20.0B48 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen2.5-0.5B-Instruct

Qwen/Qwen2.5-0.5B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

6,428,381620
500M4 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen3-4B

Qwen/Qwen3-4B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

6,283,681694
4.0B12 GB+ VRAMapache-2.0
Deployment details
time-series-forecasting

autogluon/chronos-bolt-small

autogluon/chronos-bolt-small is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,145,86961
Unknown8 GB+ VRAMapache-2.0
Deployment details
fill-mask

FacebookAI/roberta-large

FacebookAI/roberta-large is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,121,295319
Unknown8 GB+ VRAMmit
Deployment details
text-generationGated

meta-llama/Llama-3.2-1B-Instruct

meta-llama/Llama-3.2-1B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. Access approval is required on Hugging Face.

6,107,1961,611
1.0B6 GB+ VRAMllama3.2
Deployment details
text-generation

Qwen/Qwen2.5-3B-Instruct

Qwen/Qwen2.5-3B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

5,973,026564
3.0B12 GB+ VRAMother
Deployment details
time-series-forecasting

autogluon/chronos-2-small

autogluon/chronos-2-small is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,957,7756
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-ranking

cross-encoder/ms-marco-MiniLM-L4-v2

cross-encoder/ms-marco-MiniLM-L4-v2 is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,875,40229
Unknown8 GB+ VRAMapache-2.0
Deployment details