HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

752,461 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

general AI

manubc33/Qwen3.8-27B-GGUF

manubc33/Qwen3.8-27B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

3,8290
27.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

LiquidAI/LFM2.5-8B-A1B-DSpark

LiquidAI/LFM2.5-8B-A1B-DSpark is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

3,82635
8.0B24 GB+ VRAMother
Deployment details
image-text-to-text

HivenetQuant/Qwen3.8-27B-NVFP4

HivenetQuant/Qwen3.8-27B-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

3,8258
27.0B64 GB+ VRAMapache-2.0
Deployment details
text-generation

llm-jp/llm-jp-4-8b-thinking-gguf

llm-jp/llm-jp-4-8b-thinking-gguf is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,82216
8.0B8 GB+ VRAMapache-2.0
Deployment details
general AI

poolside/Laguna-S-2.1-DFlash-FP8

poolside/Laguna-S-2.1-DFlash-FP8 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,8178
Unknown8 GB+ VRAMLicense unknown
Deployment details
image-text-to-text

cyankiwi/Qwen3.8-Flash-Next-AWQ-INT4

cyankiwi/Qwen3.8-Flash-Next-AWQ-INT4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,81716
Unknown8 GB+ VRAMother
Deployment details
fill-mask

flair-bio/amplify-120m

flair-bio/amplify-120m is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,8150
Unknown8 GB+ VRAMmit
Deployment details
token-classification

stanfordnlp/stanza-hr

stanfordnlp/stanza-hr is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,8141
Unknown8 GB+ VRAMapache-2.0
Deployment details
feature-extraction

farbodtavakkoli/OTel-Embedding-109M

farbodtavakkoli/OTel-Embedding-109M is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

3,8111
Unknown4 GB+ VRAMapache-2.0
Deployment details
token-classification

cstr/fireredpunc-GGUF

cstr/fireredpunc-GGUF is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,8080
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

unsloth/Phi-3.5-mini-instruct

unsloth/Phi-3.5-mini-instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,80752
Unknown8 GB+ VRAMmit
Deployment details
sentence-similarity

sdadas/mmlw-roberta-large

sdadas/mmlw-roberta-large is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

3,80615
Unknown4 GB+ VRAMapache-2.0
Deployment details
image-to-video

aidealab/AnimeGen-I2V

aidealab/AnimeGen-I2V is a image to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,80440
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generationGated

pfnet/plamo-3-nict-2b-base

pfnet/plamo-3-nict-2b-base is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.

3,8036
2.0B8 GB+ VRAMother
Deployment details