HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

506,275 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

automatic-speech-recognition

cstr/qwen3-asr-1.7b-GGUF

cstr/qwen3-asr-1.7b-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

7,07113
1.7B4 GB+ VRAMapache-2.0
Deployment details
sentence-similarity

tardellirs/colibri-embed-ptbr

tardellirs/colibri-embed-ptbr is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

7,061105
Unknown4 GB+ VRAMgemma
Deployment details
text-to-image

ProGamerGov/qwen-360-diffusion

ProGamerGov/qwen-360-diffusion is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

7,04658
Unknown12 GB+ VRAMmit
Deployment details
visual-document-retrieval

nvidia/llama-nemotron-colembed-vl-3b-v2

nvidia/llama-nemotron-colembed-vl-3b-v2 is a visual document retrieval model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

7,03323
3.0B12 GB+ VRAMother
Deployment details
token-classification

nationaldesignstudio/rampart

nationaldesignstudio/rampart is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

7,028163
Unknown8 GB+ VRAMcc-by-4.0
Deployment details
automatic-speech-recognition

handy-computer/parakeet-tdt_ctc-1.1b-gguf

handy-computer/parakeet-tdt_ctc-1.1b-gguf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

7,0260
1.1B4 GB+ VRAMcc-by-4.0
Deployment details
text-generation

flywheel-ai/automotive

flywheel-ai/automotive is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

7,0260
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

z-lab/Qwen3.5-122B-A10B-DFlash

z-lab/Qwen3.5-122B-A10B-DFlash is a text generation model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.

7,02320
122.0B256 GB+ VRAMapache-2.0
Deployment details
text-to-image

Bedovyy/Anima-INT8

Bedovyy/Anima-INT8 is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

7,02246
Unknown12 GB+ VRAMother
Deployment details
image-text-to-text

IAAR-Shanghai/Metis-4B

IAAR-Shanghai/Metis-4B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

7,0036
4.0B12 GB+ VRAMapache-2.0
Deployment details
text-generation

SC117/Ornith-1.0-35B-MTP-APEX-GGUF

SC117/Ornith-1.0-35B-MTP-APEX-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

6,97679
35.0B32 GB+ VRAMmit
Deployment details
text-generation

OBLITERATUS/Qwen3.6-27B-OBLITERATED

OBLITERATUS/Qwen3.6-27B-OBLITERATED is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

6,960197
27.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

sudo-0x2a/Qwen3.6-27B-NVFP4-GPTQ

sudo-0x2a/Qwen3.6-27B-NVFP4-GPTQ is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

6,9577
27.0B64 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

q-future/Q-ReAlign-Pro-9B

q-future/Q-ReAlign-Pro-9B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

6,9433
9.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

microsoft/Phi-4-reasoning-vision-15B

microsoft/Phi-4-reasoning-vision-15B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.

6,925178
15.0B48 GB+ VRAMmit
Deployment details
any-to-any

onnx-community/gemma-4-E2B-it-ONNX

onnx-community/gemma-4-E2B-it-ONNX is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,90836
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

cyankiwi/Ovis2.6-30B-A3B-AWQ-4bit

cyankiwi/Ovis2.6-30B-A3B-AWQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

6,8880
30.0B24 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

handy-computer/whisper-small.en-gguf

handy-computer/whisper-small.en-gguf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,8780
Unknown8 GB+ VRAMapache-2.0
Deployment details