HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

202,823 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

general AI

mudler/KAT-Coder-V2.5-Dev-APEX-GGUF

mudler/KAT-Coder-V2.5-Dev-APEX-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,287,01355
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

ggml-org/Qwen3.8-27B-GGUF

ggml-org/Qwen3.8-27B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

1,285,91373
27.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/Qwen3.6-35B-A3B-GGUF

unsloth/Qwen3.6-35B-A3B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

1,284,1151,590
35.0B32 GB+ VRAMapache-2.0
Deployment details
sentence-similarity

Alibaba-NLP/gte-multilingual-base

Alibaba-NLP/gte-multilingual-base is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

1,279,436375
Unknown4 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

cyankiwi/Qwen3.6-27B-AWQ-INT4

cyankiwi/Qwen3.6-27B-AWQ-INT4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

1,276,965110
27.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/Qwen3.6-27B-GGUF

unsloth/Qwen3.6-27B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

1,251,265952
27.0B24 GB+ VRAMapache-2.0
Deployment details
feature-extraction

WhereIsAI/UAE-Large-V1

WhereIsAI/UAE-Large-V1 is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

1,248,561237
Unknown4 GB+ VRAMmit
Deployment details
image-text-to-text

Qwen/Qwen2-VL-7B-Instruct

Qwen/Qwen2-VL-7B-Instruct is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

1,246,5131,285
7.0B24 GB+ VRAMapache-2.0
Deployment details
time-series-forecasting

NeoQuasar/Kronos-base

NeoQuasar/Kronos-base is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,240,693267
Unknown8 GB+ VRAMmit
Deployment details
image-segmentation

CIDAS/clipseg-rd64-refined

CIDAS/clipseg-rd64-refined is a image segmentation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,219,663141
Unknown8 GB+ VRAMapache-2.0
Deployment details
time-series-forecasting

amazon/chronos-bolt-base

amazon/chronos-bolt-base is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,209,48892
Unknown8 GB+ VRAMapache-2.0
Deployment details
any-to-any

google/gemma-4-12B-it-qat-w4a16-ct

google/gemma-4-12B-it-qat-w4a16-ct is a any to any model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

1,205,24356
12.0B12 GB+ VRAMapache-2.0
Deployment details
time-series-forecasting

NeoQuasar/Kronos-small

NeoQuasar/Kronos-small is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,201,12332
Unknown8 GB+ VRAMmit
Deployment details
text-generation

nvidia/Qwen3.5-122B-A10B-NVFP4

nvidia/Qwen3.5-122B-A10B-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.

1,177,58153
122.0B256 GB+ VRAMapache-2.0
Deployment details
translation

Helsinki-NLP/opus-mt-nl-en

Helsinki-NLP/opus-mt-nl-en is a translation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,174,66810
Unknown8 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

NbAiLab/nb-wav2vec2-1b-bokmaal-v2

NbAiLab/nb-wav2vec2-1b-bokmaal-v2 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

1,172,4940
1.0B6 GB+ VRAMapache-2.0
Deployment details
text-ranking

Qwen/Qwen3-Reranker-0.6B

Qwen/Qwen3-Reranker-0.6B is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

1,171,223393
600M4 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

gagan3012/wav2vec2-xlsr-nepali

gagan3012/wav2vec2-xlsr-nepali is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,163,7208
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

ibm-research/PowerMoE-3b

ibm-research/PowerMoE-3b is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

1,162,33022
3.0B12 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen3-30B-A3B-Instruct-2507

Qwen/Qwen3-30B-A3B-Instruct-2507 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

1,161,599834
30.0B80 GB+ VRAMapache-2.0
Deployment details