HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

1,175,035 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

shoumenchougou/RWKV7-G1j-13.3B-GGUF

shoumenchougou/RWKV7-G1j-13.3B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

↓ 2,166♡ 0
13.3B12 GB+ VRAMapache-2.0
Deployment details →
general AI

mlx-community/Qwen3.5-2B-MLX-8bit

mlx-community/Qwen3.5-2B-MLX-8bit is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

↓ 2,163♡ 10
2.0B6 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

OpenVINO/Qwen3.5-9B-int8-ov

OpenVINO/Qwen3.5-9B-int8-ov is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

↓ 2,162♡ 5
9.0B12 GB+ VRAMapache-2.0
Deployment details →
image-to-image

tonera/FLUX.2-klein-4B-fp8-diffusers

tonera/FLUX.2-klein-4B-fp8-diffusers is a image to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

↓ 2,160♡ 1
4.0B12 GB+ VRAMapache-2.0
Deployment details →
general AI

snuh/hari-q3-8b

snuh/hari-q3-8b is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 2,158♡ 6
8.0B24 GB+ VRAMapache-2.0
Deployment details →
text-generation

ibm-granite/granite-4.1-3b-fp8

ibm-granite/granite-4.1-3b-fp8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 2,157♡ 8
3.0B4 GB+ VRAMapache-2.0
Deployment details →
text-generation

ibm-granite/granite-8b-code-base-4k

ibm-granite/granite-8b-code-base-4k is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 2,157♡ 32
8.0B24 GB+ VRAMapache-2.0
Deployment details →
text-generation

bartowski/allenai_SERA-8B-GGUF

bartowski/allenai_SERA-8B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 2,156♡ 0
8.0B8 GB+ VRAMapache-2.0
Deployment details →