HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

823,861 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

image-text-to-text

cyankiwi/Qwen3.5-9B-AWQ-BF16-INT8

cyankiwi/Qwen3.5-9B-AWQ-BF16-INT8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

3,5622
9.0B12 GB+ VRAMapache-2.0
Deployment details
image-to-text

PaddlePaddle/PP-OCRv6_tiny_det_onnx

PaddlePaddle/PP-OCRv6_tiny_det_onnx is a image to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,56114
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

pipenetwork/GLM-5.3-Flash-MLX-4bit

pipenetwork/GLM-5.3-Flash-MLX-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,5572
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

QuantTrio/Qwen3.5-2B-AWQ

QuantTrio/Qwen3.5-2B-AWQ is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,5524
2.0B8 GB+ VRAMapache-2.0
Deployment details
general AI

BeaverAI/Orion-26B-A4B-v1f-GGUF

BeaverAI/Orion-26B-A4B-v1f-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

3,5526
26.0B24 GB+ VRAMLicense unknown
Deployment details
general AI

AngelSlim/Qwen3-8B_eagle3

AngelSlim/Qwen3-8B_eagle3 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

3,5511
8.0B24 GB+ VRAMLicense unknown
Deployment details
text-classification

hfmlsoc/ncii-light-guard-v01

hfmlsoc/ncii-light-guard-v01 is a text classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,5512
Unknown8 GB+ VRAMmit
Deployment details
text-generation

unsloth/DeepSeek-V4-Flash-0731

unsloth/DeepSeek-V4-Flash-0731 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,54334
Unknown8 GB+ VRAMmit
Deployment details
text-generation

codelion/dhara-250m

codelion/dhara-250m is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,5384
Unknown8 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

cstr/vibevoice-asr-GGUF

cstr/vibevoice-asr-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,5377
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

bartowski/FrontisAI_Frontis-MA1-35B-GGUF

bartowski/FrontisAI_Frontis-MA1-35B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

3,5370
35.0B32 GB+ VRAMcc-by-nc-4.0
Deployment details
automatic-speech-recognition

changelinglab/PhoneticXeus

changelinglab/PhoneticXeus is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

3,53613
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

ReadyArt/Serenity-12B-GGUF

ReadyArt/Serenity-12B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

3,53327
12.0B12 GB+ VRAMapache-2.0
Deployment details