HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

264,510 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

automatic-speech-recognition

nvidia/parakeet-rnnt-0.6b

nvidia/parakeet-rnnt-0.6b is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

70,93315
600M4 GB+ VRAMcc-by-4.0
Deployment details
image-to-video

leejet/MiniMax-H3-GGUF

leejet/MiniMax-H3-GGUF is a image to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

70,80219
Unknown8 GB+ VRAMLicense unknown
Deployment details
text-to-speech

nineninesix/gepard-1.0

nineninesix/gepard-1.0 is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

70,551132
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

AtomicChat/Qwen3.5-9B-DFlash-GGUF

AtomicChat/Qwen3.5-9B-DFlash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

70,4672
9.0B8 GB+ VRAMapache-2.0
Deployment details
any-to-any

google/gemma-4-E2B

google/gemma-4-E2B is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

70,372467
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

CrashOverrideX/Quillan-Ronin

CrashOverrideX/Quillan-Ronin is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

70,2713
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

AtomicChat/Qwen3-4B-DFlash-GGUF

AtomicChat/Qwen3-4B-DFlash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

69,9230
4.0B6 GB+ VRAMmit
Deployment details
text-to-videoGated

Lightricks/LTX-2.5-Diffusers

Lightricks/LTX-2.5-Diffusers is a text to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.

69,74951
Unknown8 GB+ VRAMother
Deployment details
text-generation

AtomicChat/Qwen3.5-4B-DFlash-GGUF

AtomicChat/Qwen3.5-4B-DFlash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

69,6120
4.0B6 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

QuantTrio/gemma-4-31B-it-AWQ

QuantTrio/gemma-4-31B-it-AWQ is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

69,08214
31.0B80 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/Qwen3.8-27B-FP8

unsloth/Qwen3.8-27B-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

68,91628
27.0B64 GB+ VRAMapache-2.0
Deployment details
general AI

BAAI/seggpt-vit-large

BAAI/seggpt-vit-large is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

68,6225
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

allenai/Olmo-3-7B-Think

allenai/Olmo-3-7B-Think is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

68,575107
7.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

cyankiwi/Qwen3-VL-2B-Instruct-AWQ-4bit

cyankiwi/Qwen3-VL-2B-Instruct-AWQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

68,4631
2.0B4 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

handy-computer/gigaam-v3-e2e-rnnt-gguf

handy-computer/gigaam-v3-e2e-rnnt-gguf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

68,1332
Unknown8 GB+ VRAMmit
Deployment details
text-generation

LiquidAI/LFM2-1.2B

LiquidAI/LFM2-1.2B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

67,720366
1.2B6 GB+ VRAMother
Deployment details
image-text-to-text

twolven/Qwen3.8-27B-abliterated-AWQ-MTP

twolven/Qwen3.8-27B-abliterated-AWQ-MTP is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

67,56014
27.0B64 GB+ VRAMapache-2.0
Deployment details