HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

308,303 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

image-text-to-text

microsoft/Phi-4-reasoning-vision-15B

microsoft/Phi-4-reasoning-vision-15B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.

6,925178
15.0B48 GB+ VRAMmit
Deployment details
image-text-to-text

cyankiwi/Ovis2.6-30B-A3B-AWQ-4bit

cyankiwi/Ovis2.6-30B-A3B-AWQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

6,8880
30.0B24 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

handy-computer/whisper-small.en-gguf

handy-computer/whisper-small.en-gguf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,8780
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

bartowski/ai9stars_G9v3-3B-GGUF

bartowski/ai9stars_G9v3-3B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

6,84811
3.0B4 GB+ VRAMapache-2.0
Deployment details
text-to-video

infosave/MiniMax-H3-Turbo-cmf

infosave/MiniMax-H3-Turbo-cmf is a text to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,84312
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

QuantTrio/GLM-5.2-Int4-Int8Mix

QuantTrio/GLM-5.2-Int4-Int8Mix is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,81020
Unknown8 GB+ VRAMmit
Deployment details
automatic-speech-recognition

evewashere/cohere-transcribe-03-2026-ungated

evewashere/cohere-transcribe-03-2026-ungated is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,7950
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-to-speech

WalkingCat/Soprano-1.1-80M-GGUF

WalkingCat/Soprano-1.1-80M-GGUF is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,7900
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

mudler/Holo3-35B-A3B-APEX-GGUF

mudler/Holo3-35B-A3B-APEX-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

6,7878
35.0B32 GB+ VRAMapache-2.0
Deployment details
text-generation

z-lab/Qwen3.5-4B-DFlash

z-lab/Qwen3.5-4B-DFlash is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

6,74940
4.0B12 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

cstr/cohere-transcribe-03-2026-GGUF

cstr/cohere-transcribe-03-2026-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

6,7379
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

amd/Qwen3.5-397B-A17B-MXFP4-AttnFP8-V2

amd/Qwen3.5-397B-A17B-MXFP4-AttnFP8-V2 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.

6,7280
397.0B256 GB+ VRAMapache-2.0
Deployment details