HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

364,302 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

z-lab/Qwen3.5-122B-A10B-DFlash

z-lab/Qwen3.5-122B-A10B-DFlash is a text generation model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.

7,02320
122.0B256 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

IAAR-Shanghai/Metis-4B

IAAR-Shanghai/Metis-4B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

7,0036
4.0B12 GB+ VRAMapache-2.0
Deployment details
text-generation

SC117/Ornith-1.0-35B-MTP-APEX-GGUF

SC117/Ornith-1.0-35B-MTP-APEX-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

6,97679
35.0B32 GB+ VRAMmit
Deployment details
text-generation

OBLITERATUS/Qwen3.6-27B-OBLITERATED

OBLITERATUS/Qwen3.6-27B-OBLITERATED is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

6,960197
27.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

sudo-0x2a/Qwen3.6-27B-NVFP4-GPTQ

sudo-0x2a/Qwen3.6-27B-NVFP4-GPTQ is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

6,9577
27.0B64 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

q-future/Q-ReAlign-Pro-9B

q-future/Q-ReAlign-Pro-9B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

6,9433
9.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

microsoft/Phi-4-reasoning-vision-15B

microsoft/Phi-4-reasoning-vision-15B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.

6,925178
15.0B48 GB+ VRAMmit
Deployment details
image-text-to-text

cyankiwi/Ovis2.6-30B-A3B-AWQ-4bit

cyankiwi/Ovis2.6-30B-A3B-AWQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

6,8880
30.0B24 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

handy-computer/whisper-small.en-gguf

handy-computer/whisper-small.en-gguf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,8780
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

bartowski/ai9stars_G9v3-3B-GGUF

bartowski/ai9stars_G9v3-3B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

6,84811
3.0B4 GB+ VRAMapache-2.0
Deployment details
text-generation

allenai/Olmo-Hybrid-Instruct-SFT-7B

allenai/Olmo-Hybrid-Instruct-SFT-7B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

6,84517
7.0B24 GB+ VRAMapache-2.0
Deployment details
text-to-video

infosave/MiniMax-H3-Turbo-cmf

infosave/MiniMax-H3-Turbo-cmf is a text to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,84312
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

QuantTrio/GLM-5.2-Int4-Int8Mix

QuantTrio/GLM-5.2-Int4-Int8Mix is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,81020
Unknown8 GB+ VRAMmit
Deployment details
automatic-speech-recognition

evewashere/cohere-transcribe-03-2026-ungated

evewashere/cohere-transcribe-03-2026-ungated is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,7950
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-to-speech

WalkingCat/Soprano-1.1-80M-GGUF

WalkingCat/Soprano-1.1-80M-GGUF is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,7900
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

mudler/Holo3-35B-A3B-APEX-GGUF

mudler/Holo3-35B-A3B-APEX-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

6,7878
35.0B32 GB+ VRAMapache-2.0
Deployment details