HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

188,825 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

automatic-speech-recognition

argmaxinc/whisperkit-coreml

argmaxinc/whisperkit-coreml is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

11,619,752206
Unknown8 GB+ VRAMmit
Deployment details
text-to-speech

hexgrad/Kokoro-82M

hexgrad/Kokoro-82M is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

11,529,9256,832
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Qwen/Qwen3.5-9B

Qwen/Qwen3.5-9B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

11,389,2111,912
9.0B24 GB+ VRAMapache-2.0
Deployment details
feature-extraction

BAAI/bge-base-en-v1.5

BAAI/bge-base-en-v1.5 is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

10,859,564470
Unknown4 GB+ VRAMmit
Deployment details
general AI

unsloth/Qwen3.8-27B-GGUF

unsloth/Qwen3.8-27B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

10,675,6833,681
27.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen2.5-7B-Instruct

Qwen/Qwen2.5-7B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

10,440,5481,586
7.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

nvidia/Qwen3.6-35B-A3B-NVFP4

nvidia/Qwen3.6-35B-A3B-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

10,100,049590
35.0B80 GB+ VRAMapache-2.0
Deployment details
general AI

Bingsu/adetailer

Bingsu/adetailer is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

9,616,238764
Unknown8 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognitionGated

pyannote/speaker-diarization-3.1

pyannote/speaker-diarization-3.1 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.

9,142,1563,443
Unknown8 GB+ VRAMmit
Deployment details
time-series-forecasting

autogluon/chronos-2

autogluon/chronos-2 is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

8,879,17549
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

google/gemma-4-31B-it

google/gemma-4-31B-it is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

8,672,4403,743
31.0B80 GB+ VRAMapache-2.0
Deployment details
time-series-forecasting

amazon/chronos-bolt-small

amazon/chronos-bolt-small is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

8,650,98923
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

google/gemma-4-26B-A4B-it

google/gemma-4-26B-A4B-it is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

8,588,4681,484
26.0B64 GB+ VRAMapache-2.0
Deployment details
audio-classification

laion/clap-htsat-fused

laion/clap-htsat-fused is a audio classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

8,485,095128
Unknown8 GB+ VRAMapache-2.0
Deployment details
fill-mask

FacebookAI/roberta-base

FacebookAI/roberta-base is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

8,296,907640
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

Qwen/Qwen2.5-VL-7B-Instruct

Qwen/Qwen2.5-VL-7B-Instruct is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

7,747,6151,699
7.0B24 GB+ VRAMapache-2.0
Deployment details
feature-extraction

Qwen/Qwen3-Embedding-0.6B

Qwen/Qwen3-Embedding-0.6B is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

7,495,0421,183
600M4 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen2.5-1.5B-Instruct

Qwen/Qwen2.5-1.5B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

7,326,700820
1.5B6 GB+ VRAMapache-2.0
Deployment details
general AI

Comfy-Org/z_image_turbo

Comfy-Org/z_image_turbo is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

7,289,306870
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Qwen/Qwen3.5-4B

Qwen/Qwen3.5-4B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

7,264,578897
4.0B12 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Qwen/Qwen3.6-27B-FP8

Qwen/Qwen3.6-27B-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

7,198,015354
27.0B64 GB+ VRAMapache-2.0
Deployment details
fill-mask

distilbert/distilbert-base-uncased

distilbert/distilbert-base-uncased is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

7,138,1521,170
Unknown8 GB+ VRAMapache-2.0
Deployment details
feature-extraction

intfloat/multilingual-e5-large

intfloat/multilingual-e5-large is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

6,999,9251,248
Unknown4 GB+ VRAMmit
Deployment details