HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

639,071 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

image-text-to-text

Smoffyy/Qwen3.5-9B-Instruct-Pure-GGUF

Smoffyy/Qwen3.5-9B-Instruct-Pure-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,5154
9.0B8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

mlx-community/Qwen3.5-27B-4bit

mlx-community/Qwen3.5-27B-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

4,51149
27.0B24 GB+ VRAMapache-2.0
Deployment details
general AI

prism-ml/Bonsai-1.7B-unpacked

prism-ml/Bonsai-1.7B-unpacked is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

4,50413
1.7B6 GB+ VRAMapache-2.0
Deployment details
token-classification

chopratejas/kompress-base

chopratejas/kompress-base is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,49828
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

empero-ai/Qwythos-27B-v1

empero-ai/Qwythos-27B-v1 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

4,491167
27.0B64 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

nvidia/canary-1b-flash

nvidia/canary-1b-flash is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

4,484279
1.0B6 GB+ VRAMcc-by-4.0
Deployment details
image-text-to-text

RadixArk/GLM-5.3-Flash-NVFP4

RadixArk/GLM-5.3-Flash-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,4814
Unknown8 GB+ VRAMmit
Deployment details
feature-extraction

DreamBlooms/WeMM-Embedding-2B-GGUF

DreamBlooms/WeMM-Embedding-2B-GGUF is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

4,4767
2.0B4 GB+ VRAMother
Deployment details
fill-mask

AiLab-IMCS-UL/lv-deberta-base

AiLab-IMCS-UL/lv-deberta-base is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,4680
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

ggml-org/Qwen3-Coder-Next-GGUF

ggml-org/Qwen3-Coder-Next-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,46711
Unknown8 GB+ VRAMLicense unknown
Deployment details
text-to-image

ChrisColeTech/krea2-turbo-edit-GGUF

ChrisColeTech/krea2-turbo-edit-GGUF is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

4,46210
Unknown12 GB+ VRAMunknown
Deployment details
image-text-to-text

bartowski/OrionLLM_LRM-3.2-GGUF

bartowski/OrionLLM_LRM-3.2-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,4593
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

DavidAU/MN-DARKEST-UNIVERSE-29B-GGUF

DavidAU/MN-DARKEST-UNIVERSE-29B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

4,45886
29.0B24 GB+ VRAMapache-2.0
Deployment details
text-to-image

hoidhxd/SenseNova-U1.5-8B-GGUF-v2

hoidhxd/SenseNova-U1.5-8B-GGUF-v2 is a text to image model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

4,4583
8.0B8 GB+ VRAMLicense unknown
Deployment details