HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

873,060 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

general AI

MuyeHuang/DuplexOmni

MuyeHuang/DuplexOmni is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,0121
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Content-AI/Qwen3.5-4B

Content-AI/Qwen3.5-4B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

2,0110
4.0B12 GB+ VRAMapache-2.0
Deployment details
sentence-similarity

dnotitia/DNA-VL-STEER-2B

dnotitia/DNA-VL-STEER-2B is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,0114
2.0B8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

lightonai/LightOnOCR-2-1B-bbox-base

lightonai/LightOnOCR-2-1B-bbox-base is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

2,0104
1.0B6 GB+ VRAMapache-2.0
Deployment details
text-generation

Emiliosbs/Ben3.0-7B-Uncensored

Emiliosbs/Ben3.0-7B-Uncensored is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,0102
7.0B24 GB+ VRAMapache-2.0
Deployment details
object-detection

litert-community/yolox-nano-litert

litert-community/yolox-nano-litert is a object detection model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,0082
Unknown8 GB+ VRAMapache-2.0
Deployment details
token-classification

sallani/PrivaMesh

sallani/PrivaMesh is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,0060
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

mlx-community/Qwen3.5-4B-MTP-4bit

mlx-community/Qwen3.5-4B-MTP-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

2,0062
4.0B6 GB+ VRAMapache-2.0
Deployment details
text-to-speech

ilintar/moss-tts-gguf

ilintar/moss-tts-gguf is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,0042
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

Jackrong/Qwen3.5-9B-GLM5.1-Distill-v1

Jackrong/Qwen3.5-9B-GLM5.1-Distill-v1 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,00270
9.0B8 GB+ VRAMapache-2.0
Deployment details
text-generation

Polygl0t/Tucano2-qwen-3.7B-Instruct

Polygl0t/Tucano2-qwen-3.7B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

2,0023
3.7B12 GB+ VRAMapache-2.0
Deployment details
text-generation

WaveCut/Qwythos-9B-v2-Heretic-GGUF

WaveCut/Qwythos-9B-v2-Heretic-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,0004
9.0B8 GB+ VRAMapache-2.0
Deployment details