HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

865,057 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

sentence-similarity

sdadas/mmlw-e5-small

sdadas/mmlw-e5-small is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

2,0700
Unknown4 GB+ VRAMapache-2.0
Deployment details
fill-mask

adsabs/astroBERT

adsabs/astroBERT is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,06917
Unknown8 GB+ VRAMmit
Deployment details
text-generation

r0b0tlab/Ling-3.0-flash-NVFP4

r0b0tlab/Ling-3.0-flash-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,0683
Unknown8 GB+ VRAMmit
Deployment details
text-generation

lm-provers/QED-Nano

lm-provers/QED-Nano is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,06789
Unknown8 GB+ VRAMapache-2.0
Deployment details
token-classification

KRLabsOrg/lettucedect-v2-mmbert-base

KRLabsOrg/lettucedect-v2-mmbert-base is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,0674
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

kaptaan45/QaptaanLM-0.75B

kaptaan45/QaptaanLM-0.75B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

2,0660
750M4 GB+ VRAMapache-2.0
Deployment details
text-generation

litert-community/Qwen3.5-4B

litert-community/Qwen3.5-4B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

2,0662
4.0B12 GB+ VRAMapache-2.0
Deployment details
text-generation

Gryphe/Gemma-4-31B-StyleTune

Gryphe/Gemma-4-31B-StyleTune is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

2,06488
31.0B80 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

cstr/mimo-asr-GGUF

cstr/mimo-asr-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,0638
Unknown8 GB+ VRAMmit
Deployment details
text-generation

dnotitia/Qwen3-0.6B

dnotitia/Qwen3-0.6B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

2,0580
600M4 GB+ VRAMapache-2.0
Deployment details
audio-classification

funasr/campplus

funasr/campplus is a audio classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,05726
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

MayaKD/qwen2-vl-audio

MayaKD/qwen2-vl-audio is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,0570
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-to-speech

g-group-ai-lab/gwen-tts-0.6B

g-group-ai-lab/gwen-tts-0.6B is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

2,05527
600M4 GB+ VRAMmit
Deployment details
text-generation

incoai/Muse-Glimmer-30B-DFlash2

incoai/Muse-Glimmer-30B-DFlash2 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

2,0559
30.0B80 GB+ VRAMapache-2.0
Deployment details
text-generation

furiosa-ai/Qwen3-32B-FP8

furiosa-ai/Qwen3-32B-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

2,0540
32.0B80 GB+ VRAMapache-2.0
Deployment details