HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

390,301 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

sentence-similarity

lightonai/DenseOn

lightonai/DenseOn is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

11,28135
Unknown4 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

handy-computer/parakeet-rnnt-1.1b-gguf

handy-computer/parakeet-rnnt-1.1b-gguf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

11,2661
1.1B4 GB+ VRAMcc-by-4.0
Deployment details
visual-document-retrieval

TomoroAI/tomoro-colqwen3-embed-4b

TomoroAI/tomoro-colqwen3-embed-4b is a visual document retrieval model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

11,23233
4.0B12 GB+ VRAMapache-2.0
Deployment details
feature-extraction

marcoyang/spear-xlarge-speech-audio

marcoyang/spear-xlarge-speech-audio is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

11,2287
Unknown4 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

cyankiwi/Qwen3.5-397B-A17B-AWQ-4bit

cyankiwi/Qwen3.5-397B-A17B-AWQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.

11,1913
397.0B256 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

mlx-community/gemma-4-E4B-it-qat-mobile

mlx-community/gemma-4-E4B-it-qat-mobile is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

11,1850
Unknown8 GB+ VRAMLicense unknown
Deployment details
image-to-image

unsloth/FLUX.2-klein-9B

unsloth/FLUX.2-klein-9B is a image to image model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

11,1506
9.0B24 GB+ VRAMother
Deployment details
image-text-to-text

Jackrong/Qwopus3.5-9B-Coder-GGUF

Jackrong/Qwopus3.5-9B-Coder-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

11,127339
9.0B8 GB+ VRAMapache-2.0
Deployment details
text-to-image

LeFeujitif/sandbox

LeFeujitif/sandbox is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

11,0913
Unknown12 GB+ VRAMLicense unknown
Deployment details
general AI

ReadyArt/Serenity-27B-GGUF

ReadyArt/Serenity-27B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

11,0889
27.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

zai-org/GLM-5.3-Flash-BF16

zai-org/GLM-5.3-Flash-BF16 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

11,07558
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

dealignai/Bonsai-27b-Ternary-CRACK-GGUF

dealignai/Bonsai-27b-Ternary-CRACK-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

11,07324
27.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

empero-ai/Qwable-9B-Claude-Fable-5

empero-ai/Qwable-9B-Claude-Fable-5 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

11,069118
9.0B24 GB+ VRAMapache-2.0
Deployment details
general AI

jagat334433/beru_custom

jagat334433/beru_custom is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

11,0599
Unknown8 GB+ VRAMmit
Deployment details
text-to-image

unsloth/Z-Image-Turbo

unsloth/Z-Image-Turbo is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

11,0505
Unknown12 GB+ VRAMapache-2.0
Deployment details