HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

975,049 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

image-text-to-text

coder3101/Qwen3.8-27B-heretic

coder3101/Qwen3.8-27B-heretic is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.

↓ 2,531♡ 4
27.0B64 GB+ VRAMapache-2.0
Deployment details →
text-generation

sarvamai/sarvam-30b-fp8

sarvamai/sarvam-30b-fp8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.

↓ 2,529♡ 14
30.0B80 GB+ VRAMapache-2.0
Deployment details →
text-generation

mradermacher/RIFA-FLASH-1.7B-i1-GGUF

mradermacher/RIFA-FLASH-1.7B-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 2,525♡ 0
1.7B4 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

AlexHung29629/gemma-4-E2B

AlexHung29629/gemma-4-E2B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 2,522♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →
general AI

nightmedia/Qwen3.5-4B-mxfp4-mlx

nightmedia/Qwen3.5-4B-mxfp4-mlx is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

↓ 2,517♡ 0
4.0B12 GB+ VRAMapache-2.0
Deployment details →
image-segmentation

Roboflow/rf-detr-segmentation

Roboflow/rf-detr-segmentation is a image segmentation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 2,514♡ 25
Unknown8 GB+ VRAMapache-2.0
Deployment details →
general AI

Kecven/Qwen3.8-27B-MTPLX-Q4

Kecven/Qwen3.8-27B-MTPLX-Q4 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 2,513♡ 4
27.0B24 GB+ VRAMapache-2.0
Deployment details →
text-generation

nm-testing/convert_modelopt_nvfp4-e2e

nm-testing/convert_modelopt_nvfp4-e2e is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 2,513♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →
sentence-similarity

NbAiLab/nb-sbert-v2-large

NbAiLab/nb-sbert-v2-large is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 2,510♡ 3
Unknown4 GB+ VRAMapache-2.0
Deployment details →
image-to-video

QuantStack/LTX-2-GGUF

QuantStack/LTX-2-GGUF is a image to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 2,509♡ 81
Unknown8 GB+ VRAMother
Deployment details →