HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

398,301 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

arcee-ai/Trinity-Nano-Preview

arcee-ai/Trinity-Nano-Preview is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

10,35178
Unknown8 GB+ VRAMother
Deployment details
token-classification

stanfordnlp/stanza-en

stanfordnlp/stanza-en is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

10,34417
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

cyankiwi/Ornith-1.0-9B-AWQ-INT4

cyankiwi/Ornith-1.0-9B-AWQ-INT4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

10,3331
9.0B8 GB+ VRAMmit
Deployment details
text-generation

openbmb/MiniCPM5-1B-Base

openbmb/MiniCPM5-1B-Base is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

10,33114
1.0B6 GB+ VRAMapache-2.0
Deployment details
general AI

tencent/Hy-MT2-1.8B-FP8

tencent/Hy-MT2-1.8B-FP8 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

10,32814
1.8B6 GB+ VRAMapache-2.0
Deployment details
text-generation

abenzerps/Apodex-1.1-mini-GGUF

abenzerps/Apodex-1.1-mini-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

10,29912
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

amd/gpt-oss-120b-w-mxfp4-a-fp8

amd/gpt-oss-120b-w-mxfp4-a-fp8 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.

10,29710
120.0B256 GB+ VRAMapache-2.0
Deployment details
feature-extraction

tencent/WeMM-Embedding-2B

tencent/WeMM-Embedding-2B is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

10,27883
2.0B8 GB+ VRAMother
Deployment details
text-generation

electricpipelines/Qwen3-4B-GGUF

electricpipelines/Qwen3-4B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

10,2530
4.0B6 GB+ VRAMapache-2.0
Deployment details
general AI

ibm-granite/granitelib-rag-r1.0

ibm-granite/granitelib-rag-r1.0 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

10,23748
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

LiquidAI/LFM2-700M-GGUF

LiquidAI/LFM2-700M-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

10,23554
Unknown8 GB+ VRAMother
Deployment details
text-generation

Avifenesh/Qwen3.8-27B-NVFP4-MTP-GGUF

Avifenesh/Qwen3.8-27B-NVFP4-MTP-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

10,2350
27.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

lactroiii/Qwen3.8-27B-Uncensored-GGUF

lactroiii/Qwen3.8-27B-Uncensored-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

10,2172
27.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Vontra/Qwen3.8-Flash-Next-MLX-4bit

Vontra/Qwen3.8-Flash-Next-MLX-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

10,2155
Unknown8 GB+ VRAMother
Deployment details
general AI

Alissonerdx/EditAnything

Alissonerdx/EditAnything is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

10,20995
Unknown8 GB+ VRAMapache-2.0
Deployment details
depth-estimation

mudler/depth-anything.cpp-gguf

mudler/depth-anything.cpp-gguf is a depth estimation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

10,18312
Unknown8 GB+ VRAMapache-2.0
Deployment details
token-classification

bardsai/eu-pii-anonimization-multilang

bardsai/eu-pii-anonimization-multilang is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

10,13918
Unknown8 GB+ VRAMapache-2.0
Deployment details