HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

863,057 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

drawais/Qwen3-Embedding-4B-AWQ-INT4

drawais/Qwen3-Embedding-4B-AWQ-INT4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

3,1335
4.0B6 GB+ VRAMapache-2.0
Deployment details
any-to-any

unsloth/gemma-4-E4B-it-qat-w4a16

unsloth/gemma-4-E4B-it-qat-w4a16 is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,1313
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

EleutherAI/pythia-160m-seed4

EleutherAI/pythia-160m-seed4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,1281
Unknown8 GB+ VRAMapache-2.0
Deployment details
sentence-similarity

infly/inf-retriever-v1

infly/inf-retriever-v1 is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

3,12845
Unknown4 GB+ VRAMapache-2.0
Deployment details
text-generation

LiquidAI/LFM2-2.6B-Transcript-GGUF

LiquidAI/LFM2-2.6B-Transcript-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

3,12624
2.6B4 GB+ VRAMother
Deployment details
general AI

Jamphus/PinkCherry_NSFW_LTX23

Jamphus/PinkCherry_NSFW_LTX23 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,1222
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

CrossNow/Qwen3.8-27B-GGUF

CrossNow/Qwen3.8-27B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

3,1210
27.0B24 GB+ VRAMapache-2.0
Deployment details
general AI

HeartMuLa/HeartCodec-oss-20260123

HeartMuLa/HeartCodec-oss-20260123 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,12010
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

amd/gpt-oss-20b-WFP8-AFP8-KVFP8

amd/gpt-oss-20b-WFP8-AFP8-KVFP8 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.

3,1200
20.0B48 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

FastFlowLM/Qwen3.6-35B-A3B-NPU2

FastFlowLM/Qwen3.6-35B-A3B-NPU2 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

3,12015
35.0B80 GB+ VRAMapache-2.0
Deployment details
text-generation

ai9stars/G9v3-3B

ai9stars/G9v3-3B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

3,11958
3.0B12 GB+ VRAMapache-2.0
Deployment details
text-to-audio

ACE-Step/acestep-v15-base

ACE-Step/acestep-v15-base is a text to audio model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,11367
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

tvall43/Qwen3.5-4B-heretic-v2

tvall43/Qwen3.5-4B-heretic-v2 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

3,1126
4.0B12 GB+ VRAMapache-2.0
Deployment details