HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

202,823 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

image-text-to-text

ggml-org/Qwen3.8-27B-GGUF

ggml-org/Qwen3.8-27B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

1,285,91373
27.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/Qwen3.6-35B-A3B-GGUF

unsloth/Qwen3.6-35B-A3B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

1,284,1151,590
35.0B32 GB+ VRAMapache-2.0
Deployment details
sentence-similarity

Alibaba-NLP/gte-multilingual-base

Alibaba-NLP/gte-multilingual-base is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

1,279,436375
Unknown4 GB+ VRAMapache-2.0
Deployment details
text-generationGated

meta-llama/Llama-3.2-1B

meta-llama/Llama-3.2-1B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. Access approval is required on Hugging Face.

1,277,1842,579
1.0B6 GB+ VRAMllama3.2
Deployment details
image-text-to-text

cyankiwi/Qwen3.6-27B-AWQ-INT4

cyankiwi/Qwen3.6-27B-AWQ-INT4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

1,276,965110
27.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

MiniMaxAI/MiniMax-M2.7

MiniMaxAI/MiniMax-M2.7 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,267,5971,246
Unknown8 GB+ VRAMother
Deployment details
image-text-to-text

datalab-to/surya-ocr-2

datalab-to/surya-ocr-2 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,258,621104
Unknown8 GB+ VRAMopenrail
Deployment details
image-text-to-text

gaunernst/gemma-3-27b-it-int4-awq

gaunernst/gemma-3-27b-it-int4-awq is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

1,254,75240
27.0B24 GB+ VRAMgemma
Deployment details
image-text-to-text

unsloth/Qwen3.6-27B-GGUF

unsloth/Qwen3.6-27B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

1,251,265952
27.0B24 GB+ VRAMapache-2.0
Deployment details
feature-extraction

WhereIsAI/UAE-Large-V1

WhereIsAI/UAE-Large-V1 is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

1,248,561237
Unknown4 GB+ VRAMmit
Deployment details
image-text-to-text

Qwen/Qwen2-VL-7B-Instruct

Qwen/Qwen2-VL-7B-Instruct is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

1,246,5131,285
7.0B24 GB+ VRAMapache-2.0
Deployment details
time-series-forecasting

NeoQuasar/Kronos-base

NeoQuasar/Kronos-base is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,240,693267
Unknown8 GB+ VRAMmit
Deployment details
automatic-speech-recognition

classla/wav2vec2-xls-r-parlaspeech-hr

classla/wav2vec2-xls-r-parlaspeech-hr is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,233,5703
Unknown8 GB+ VRAMLicense unknown
Deployment details
general AI

zhihan1996/DNABERT-2-117M

zhihan1996/DNABERT-2-117M is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,221,039112
Unknown8 GB+ VRAMLicense unknown
Deployment details
image-segmentation

CIDAS/clipseg-rd64-refined

CIDAS/clipseg-rd64-refined is a image segmentation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,219,663141
Unknown8 GB+ VRAMapache-2.0
Deployment details
time-series-forecasting

amazon/chronos-bolt-base

amazon/chronos-bolt-base is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,209,48892
Unknown8 GB+ VRAMapache-2.0
Deployment details
any-to-any

google/gemma-4-12B-it-qat-w4a16-ct

google/gemma-4-12B-it-qat-w4a16-ct is a any to any model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

1,205,24356
12.0B12 GB+ VRAMapache-2.0
Deployment details
time-series-forecasting

NeoQuasar/Kronos-small

NeoQuasar/Kronos-small is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,201,12332
Unknown8 GB+ VRAMmit
Deployment details
automatic-speech-recognition

nvidia/parakeet-ctc-1.1b

nvidia/parakeet-ctc-1.1b is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

1,198,02958
1.1B4 GB+ VRAMcc-by-4.0
Deployment details
image-to-video

Lightricks/LTX-2.3

Lightricks/LTX-2.3 is a image to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,181,8281,876
Unknown8 GB+ VRAMother
Deployment details
text-generation

nvidia/Qwen3.5-122B-A10B-NVFP4

nvidia/Qwen3.5-122B-A10B-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.

1,177,58153
122.0B256 GB+ VRAMapache-2.0
Deployment details