HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

196,823 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

HuggingFaceTB/SmolLM2-135M

HuggingFaceTB/SmolLM2-135M is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,550,777230
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

colbert-ir/colbertv2.0

colbert-ir/colbertv2.0 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,533,627366
Unknown8 GB+ VRAMmit
Deployment details
zero-shot-image-classification

openai/clip-vit-large-patch14-336

openai/clip-vit-large-patch14-336 is a zero shot image classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,528,177309
Unknown8 GB+ VRAMLicense unknown
Deployment details
text-classification

daekeun-ml/koelectra-small-v3-nsmc

daekeun-ml/koelectra-small-v3-nsmc is a text classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,527,0797
Unknown8 GB+ VRAMmit
Deployment details
text-generationGated

google/gemma-3-1b-it

google/gemma-3-1b-it is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. Access approval is required on Hugging Face.

2,492,2021,137
1.0B6 GB+ VRAMgemma
Deployment details
text-ranking

Qwen/Qwen3-Reranker-4B

Qwen/Qwen3-Reranker-4B is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

2,485,166155
4.0B12 GB+ VRAMapache-2.0
Deployment details
fill-mask

biohub/ESMC-6B

biohub/ESMC-6B is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 16 GB. It is publicly listed on Hugging Face.

2,467,39831
6.0B16 GB+ VRAMmit
Deployment details
general AI

mistralai/Mistral-7B-Instruct-v0.3

mistralai/Mistral-7B-Instruct-v0.3 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,461,0002,839
7.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Qwen/Qwen3-VL-8B-Instruct-FP8

Qwen/Qwen3-VL-8B-Instruct-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,453,73282
8.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen2.5-Coder-7B-Instruct

Qwen/Qwen2.5-Coder-7B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,440,111794
7.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/Qwen3.6-27B-NVFP4

unsloth/Qwen3.6-27B-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

2,432,022278
27.0B64 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Qwen/Qwen3.5-0.8B

Qwen/Qwen3.5-0.8B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

2,412,896698
800M4 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

deepseek-ai/DeepSeek-OCR

deepseek-ai/DeepSeek-OCR is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,389,6393,351
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

moonshotai/Kimi-K3

moonshotai/Kimi-K3 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,380,44811,241
Unknown8 GB+ VRAMother
Deployment details
time-series-forecasting

autogluon/chronos-bolt-tiny

autogluon/chronos-bolt-tiny is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,378,94413
Unknown8 GB+ VRAMapache-2.0
Deployment details
feature-extraction

Xenova/bge-base-en-v1.5

Xenova/bge-base-en-v1.5 is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

2,350,36010
Unknown4 GB+ VRAMmit
Deployment details
automatic-speech-recognition

Systran/faster-whisper-small

Systran/faster-whisper-small is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,348,23747
Unknown8 GB+ VRAMmit
Deployment details
general AI

Comfy-Org/Qwen-Image_ComfyUI

Comfy-Org/Qwen-Image_ComfyUI is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,327,318478
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-to-speech

audio-cpp/audio.cpp-gguf

audio-cpp/audio.cpp-gguf is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,313,604107
Unknown8 GB+ VRAMother
Deployment details
automatic-speech-recognition

mistralai/Voxtral-Mini-4B-Realtime-2602

mistralai/Voxtral-Mini-4B-Realtime-2602 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

2,308,239976
4.0B12 GB+ VRAMapache-2.0
Deployment details