HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

538,273 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

image-text-to-text

stepfun-ai/Step-3.7-Flash-GGUF

stepfun-ai/Step-3.7-Flash-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,159172
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

drawais/Qwen3-Reranker-0.6B-AWQ-INT4

drawais/Qwen3-Reranker-0.6B-AWQ-INT4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

5,1580
600M4 GB+ VRAMapache-2.0
Deployment details
text-generation

reaperdoesntknow/TameForCasualLM

reaperdoesntknow/TameForCasualLM is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,1550
Unknown8 GB+ VRAMapache-2.0
Deployment details
sentence-similarity

LiquidAI/LFM2.5-Embedding-350M-GGUF

LiquidAI/LFM2.5-Embedding-350M-GGUF is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

5,15238
Unknown4 GB+ VRAMother
Deployment details
text-generation

z-lab/Muse-Glimmer-30B-DFlash2-GGUF

z-lab/Muse-Glimmer-30B-DFlash2-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

5,13712
30.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

rostlabs/rost-1b-instruct

rostlabs/rost-1b-instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

5,1321
1.0B6 GB+ VRAMcc-by-nc-4.0
Deployment details
image-text-to-text

QQZ2026/Qwen3.8-27B-NVFP4-Q5K-no-MTP-GGUF

QQZ2026/Qwen3.8-27B-NVFP4-Q5K-no-MTP-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

5,1266
27.0B24 GB+ VRAMLicense unknown
Deployment details
text-generation

RadixArk/Muse-Glimmer-q4k-dynamic-MLX

RadixArk/Muse-Glimmer-q4k-dynamic-MLX is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,1153
Unknown8 GB+ VRAMLicense unknown
Deployment details
general AI

nvidia/NV-Raw2insights-MRI

nvidia/NV-Raw2insights-MRI is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

5,11411
Unknown8 GB+ VRAMother
Deployment details
image-text-to-text

unsloth/Qwen3.5-27B

unsloth/Qwen3.5-27B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

5,11416
27.0B64 GB+ VRAMapache-2.0
Deployment details
text-to-speechGated

neuphonic/neutts-air

neuphonic/neutts-air is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.

5,113889
Unknown8 GB+ VRAMapache-2.0
Deployment details