HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

202,823 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

automatic-speech-recognition

mlx-community/parakeet-tdt-0.6b-v2

mlx-community/parakeet-tdt-0.6b-v2 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

1,625,81047
600M4 GB+ VRAMcc-by-4.0
Deployment details
automatic-speech-recognition

openai/whisper-tiny

openai/whisper-tiny is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,624,155443
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

unsloth/Qwen3.5-9B-GGUF

unsloth/Qwen3.5-9B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,618,479889
9.0B8 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

theainerd/Wav2Vec2-large-xlsr-hindi

theainerd/Wav2Vec2-large-xlsr-hindi is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,618,05313
Unknown8 GB+ VRAMLicense unknown
Deployment details
image-text-to-text

vikhyatk/moondream2

vikhyatk/moondream2 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,615,1841,435
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

nvidia/Cosmos3-Edge

nvidia/Cosmos3-Edge is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,614,126194
Unknown8 GB+ VRAMother
Deployment details
image-text-to-text

Qwen/Qwen3.5-122B-A10B-FP8

Qwen/Qwen3.5-122B-A10B-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.

1,612,865115
122.0B256 GB+ VRAMapache-2.0
Deployment details
general AI

google/mobilebert-uncased

google/mobilebert-uncased is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,610,04175
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

Qwen/Qwen2.5-Coder-32B-Instruct

Qwen/Qwen2.5-Coder-32B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

1,603,9912,128
32.0B80 GB+ VRAMapache-2.0
Deployment details
general AI

Comfy-Org/Qwen-Image-Edit_ComfyUI

Comfy-Org/Qwen-Image-Edit_ComfyUI is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,594,223471
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

deepseek-ai/DeepSeek-V3.2

deepseek-ai/DeepSeek-V3.2 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,579,5701,474
Unknown8 GB+ VRAMmit
Deployment details
automatic-speech-recognition

nguyenvulebinh/wav2vec2-base-vi-vlsp2020

nguyenvulebinh/wav2vec2-base-vi-vlsp2020 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,578,0212
Unknown8 GB+ VRAMcc-by-nc-4.0
Deployment details
text-to-image

stable-diffusion-v1-5/stable-diffusion-v1-5

stable-diffusion-v1-5/stable-diffusion-v1-5 is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

1,563,4251,266
Unknown12 GB+ VRAMcreativeml-openrail-m
Deployment details
automatic-speech-recognition

handy-computer/parakeet-unified-en-0.6b-gguf

handy-computer/parakeet-unified-en-0.6b-gguf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

1,552,1966
600M4 GB+ VRAMcc-by-4.0
Deployment details
text-generation

Qwen/Qwen3-8B-AWQ

Qwen/Qwen3-8B-AWQ is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

1,547,72854
8.0B24 GB+ VRAMapache-2.0
Deployment details
general AI

google/flan-t5-base

google/flan-t5-base is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,535,3211,091
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generationGated

meta-llama/Llama-3.2-3B-Instruct

meta-llama/Llama-3.2-3B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. Access approval is required on Hugging Face.

1,529,7512,535
3.0B12 GB+ VRAMllama3.2
Deployment details
automatic-speech-recognition

imvladikon/wav2vec2-xls-r-300m-hebrew

imvladikon/wav2vec2-xls-r-300m-hebrew is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,528,2356
Unknown8 GB+ VRAMLicense unknown
Deployment details
text-generation

farbodtavakkoli/OTel-LLM-E4B-IT

farbodtavakkoli/OTel-LLM-E4B-IT is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,525,1600
Unknown8 GB+ VRAMapache-2.0
Deployment details
audio-classification

mudler/ced-gguf

mudler/ced-gguf is a audio classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,524,9253
Unknown8 GB+ VRAMapache-2.0
Deployment details
zero-shot-image-classification

google/siglip2-base-patch16-224

google/siglip2-base-patch16-224 is a zero shot image classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,516,333134
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

rhasspy/faster-whisper-tiny-int8

rhasspy/faster-whisper-tiny-int8 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,492,2407
Unknown8 GB+ VRAMmit
Deployment details