HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

186,826 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

image-text-to-text

Qwen/Qwen3.6-35B-A3B-FP8

Qwen/Qwen3.6-35B-A3B-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

12,372,077376
35.0B80 GB+ VRAMapache-2.0
Deployment details
sentence-similarity

intfloat/multilingual-e5-small

intfloat/multilingual-e5-small is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

12,313,873399
Unknown4 GB+ VRAMmit
Deployment details
image-classification

timm/efficientnet_b3.ra2_in1k

timm/efficientnet_b3.ra2_in1k is a image classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

12,216,3187
Unknown8 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

argmaxinc/whisperkit-coreml

argmaxinc/whisperkit-coreml is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

11,619,752206
Unknown8 GB+ VRAMmit
Deployment details
text-to-speech

hexgrad/Kokoro-82M

hexgrad/Kokoro-82M is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

11,529,9256,832
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Qwen/Qwen3.5-9B

Qwen/Qwen3.5-9B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

11,389,2111,912
9.0B24 GB+ VRAMapache-2.0
Deployment details
feature-extraction

BAAI/bge-base-en-v1.5

BAAI/bge-base-en-v1.5 is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

10,859,564470
Unknown4 GB+ VRAMmit
Deployment details
general AI

unsloth/Qwen3.8-27B-GGUF

unsloth/Qwen3.8-27B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

10,675,6833,681
27.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen2.5-7B-Instruct

Qwen/Qwen2.5-7B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

10,440,5481,586
7.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

nvidia/Qwen3.6-35B-A3B-NVFP4

nvidia/Qwen3.6-35B-A3B-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

10,100,049590
35.0B80 GB+ VRAMapache-2.0
Deployment details
general AI

Bingsu/adetailer

Bingsu/adetailer is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

9,616,238764
Unknown8 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognitionGated

pyannote/speaker-diarization-3.1

pyannote/speaker-diarization-3.1 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.

9,142,1563,443
Unknown8 GB+ VRAMmit
Deployment details
time-series-forecasting

autogluon/chronos-2

autogluon/chronos-2 is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

8,879,17549
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

facebook/opt-125m

facebook/opt-125m is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

8,820,283296
Unknown8 GB+ VRAMother
Deployment details
image-text-to-text

google/gemma-4-31B-it

google/gemma-4-31B-it is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

8,672,4403,743
31.0B80 GB+ VRAMapache-2.0
Deployment details
time-series-forecasting

amazon/chronos-bolt-small

amazon/chronos-bolt-small is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

8,650,98923
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

google/gemma-4-26B-A4B-it

google/gemma-4-26B-A4B-it is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

8,588,4681,484
26.0B64 GB+ VRAMapache-2.0
Deployment details
audio-classification

laion/clap-htsat-fused

laion/clap-htsat-fused is a audio classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

8,485,095128
Unknown8 GB+ VRAMapache-2.0
Deployment details
fill-mask

FacebookAI/roberta-base

FacebookAI/roberta-base is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

8,296,907640
Unknown8 GB+ VRAMmit
Deployment details
general AI

facebook/contriever

facebook/contriever is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

7,936,73497
Unknown8 GB+ VRAMLicense unknown
Deployment details
image-text-to-text

Qwen/Qwen2.5-VL-7B-Instruct

Qwen/Qwen2.5-VL-7B-Instruct is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

7,747,6151,699
7.0B24 GB+ VRAMapache-2.0
Deployment details
feature-extraction

Qwen/Qwen3-Embedding-0.6B

Qwen/Qwen3-Embedding-0.6B is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

7,495,0421,183
600M4 GB+ VRAMapache-2.0
Deployment details