HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

901,055 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

nvidia/DeepSeek-V4-Flash-0731-NVFP4

nvidia/DeepSeek-V4-Flash-0731-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,92930
Unknown8 GB+ VRAMmit
Deployment details
image-text-to-text

6block/Qwen3.8-27B-GGUF

6block/Qwen3.8-27B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

1,9280
27.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

Atomic-Germ/DynaGuard-8B-NPU2

Atomic-Germ/DynaGuard-8B-NPU2 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,9280
8.0B8 GB+ VRAMapache-2.0
Deployment details
automatic-speech-recognition

Trelis/whisper-hinglish-preview

Trelis/whisper-hinglish-preview is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,9269
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

agustindxm/qwen-coder-jailbreak

agustindxm/qwen-coder-jailbreak is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,9265
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

inclusionAI/Ling-3.0-flash-base

inclusionAI/Ling-3.0-flash-base is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,9266
Unknown8 GB+ VRAMmit
Deployment details
text-generation

EleutherAI/pythia-410m-seed8

EleutherAI/pythia-410m-seed8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,9250
Unknown8 GB+ VRAMapache-2.0
Deployment details
text-generation

johninthepool/Qwen3.8-27B-MTPLX-bf16

johninthepool/Qwen3.8-27B-MTPLX-bf16 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

1,9242
27.0B64 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

mlx-community/Qwen3.6-35B-A3B-bf16

mlx-community/Qwen3.6-35B-A3B-bf16 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

1,92311
35.0B80 GB+ VRAMapache-2.0
Deployment details
text-generation

Ttimms/zaya1-8b-nvfp4-w4a4

Ttimms/zaya1-8b-nvfp4-w4a4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

1,9220
8.0B24 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

mlx-community/DeepSeek-OCR-8bit

mlx-community/DeepSeek-OCR-8bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,92138
Unknown8 GB+ VRAMmit
Deployment details
text-generation

mlx-community/GLM-4.7-Flash-6bit

mlx-community/GLM-4.7-Flash-6bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

1,9207
Unknown8 GB+ VRAMmit
Deployment details