HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

478,281 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

image-text-to-text

AtomicChat/gemma-4-E4B-it-GGUF

AtomicChat/gemma-4-E4B-it-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,8601
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

tencent/Hy4-preview-FP8

tencent/Hy4-preview-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,85922
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

SZLHOLDINGS/chaski

SZLHOLDINGS/chaski is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,8510
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

microsoft/UserLM-8b

microsoft/UserLM-8b is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,849388
8.0B24 GB+ VRAMmit
Deployment details
text-generation

osunlp/QUEST-35B-RL

osunlp/QUEST-35B-RL is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

2,84234
35.0B80 GB+ VRAMapache-2.0
Deployment details
text-generation

Tdamre/MiniCPM5-1B-litert-lm

Tdamre/MiniCPM5-1B-litert-lm is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

2,8423
1.0B6 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

lmstudio-community/Qwen3.5-2B-MLX-8bit

lmstudio-community/Qwen3.5-2B-MLX-8bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

2,8360
2.0B6 GB+ VRAMapache-2.0
Deployment details
text-generation

logic65/Qwen3.8-Whittle-tri-14.7B

logic65/Qwen3.8-Whittle-tri-14.7B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

2,8321
14.7B12 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

handy-computer/whisper-base.en-gguf

handy-computer/whisper-base.en-gguf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,8310
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

mobilint/Qwen3-VL-8B-Instruct-Batch16

mobilint/Qwen3-VL-8B-Instruct-Batch16 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,8270
8.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

TrevorJS/gemma-4-E4B-it-uncensored

TrevorJS/gemma-4-E4B-it-uncensored is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,82237
Unknown8 GB+ VRAMapache-2.0
Deployment details