HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

1,608,988 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

sentence-similarity

michaelfeil/embeddinggemma-300m

michaelfeil/embeddinggemma-300m is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 1,909♡ 0
Unknown4 GB+ VRAMgemma
Deployment details →
text-generation

meshllm/GLM-4.7-Flash-MTP-GGUF

meshllm/GLM-4.7-Flash-MTP-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,908♡ 0
Unknown8 GB+ VRAMmit
Deployment details →
text-generation

BioTorch/UI_Llama_3_2_1B

BioTorch/UI_Llama_3_2_1B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

↓ 1,908♡ 1
1.0B6 GB+ VRAMapache-2.0
Deployment details →
translation

Xenova/opus-mt-sv-en

Xenova/opus-mt-sv-en is a translation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,907♡ 1
Unknown8 GB+ VRAMLicense unknown
Deployment details →
text-generation

m8than/gemma-2-9b-it

m8than/gemma-2-9b-it is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 1,906♡ 0
9.0B24 GB+ VRAMgemma
Deployment details →
general AI

wop/littlechat-50M

wop/littlechat-50M is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,906♡ 2
Unknown8 GB+ VRAMLicense unknown
Deployment details →
general AI

mradermacher/zeta-2.1-GGUF

mradermacher/zeta-2.1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,905♡ 6
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-generation

miifanboy/Spark-X2.5-4B-i1-GGUF

miifanboy/Spark-X2.5-4B-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

↓ 1,905♡ 1
4.0B6 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

AxionML/Qwen3.5-0.8B-NVFP4

AxionML/Qwen3.5-0.8B-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 1,904♡ 1
800M4 GB+ VRAMapache-2.0
Deployment details →
token-classification

stanfordnlp/stanza-fr

stanfordnlp/stanza-fr is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,904♡ 2
Unknown8 GB+ VRAMapache-2.0
Deployment details →
automatic-speech-recognition

cstr/voxtral-mini-3b-2507-GGUF

cstr/voxtral-mini-3b-2507-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 1,902♡ 0
3.0B4 GB+ VRAMapache-2.0
Deployment details →