HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

561,074 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

text-generation

mradermacher/RIFA-FLASH-1.7B-i1-GGUF

mradermacher/RIFA-FLASH-1.7B-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

2,5250
1.7B4 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

AlexHung29629/gemma-4-E2B

AlexHung29629/gemma-4-E2B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,5220
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-segmentation

Roboflow/rf-detr-segmentation

Roboflow/rf-detr-segmentation is a image segmentation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,51425
Unknown8 GB+ VRAMapache-2.0
Deployment details
general AI

Kecven/Qwen3.8-27B-MTPLX-Q4

Kecven/Qwen3.8-27B-MTPLX-Q4 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,5134
27.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

nm-testing/convert_modelopt_nvfp4-e2e

nm-testing/convert_modelopt_nvfp4-e2e is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,5130
Unknown8 GB+ VRAMapache-2.0
Deployment details
sentence-similarity

NbAiLab/nb-sbert-v2-large

NbAiLab/nb-sbert-v2-large is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

2,5103
Unknown4 GB+ VRAMapache-2.0
Deployment details
text-generation

marlalabsAI/Qwen3.5-4B-SX8

marlalabsAI/Qwen3.5-4B-SX8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.

2,5070
4.0B6 GB+ VRAMapache-2.0
Deployment details
feature-extraction

codefuse-ai/F2LLM-v2-8B

codefuse-ai/F2LLM-v2-8B is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,50510
8.0B24 GB+ VRAMapache-2.0
Deployment details
text-generation

reaperdoesntknow/CasualSwarms

reaperdoesntknow/CasualSwarms is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,5002
Unknown8 GB+ VRAMapache-2.0
Deployment details
image-text-to-text

byteshape/Qwen3.5-9B-GGUF

byteshape/Qwen3.5-9B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

2,49837
9.0B8 GB+ VRAMapache-2.0
Deployment details
text-generation

BayesRL/Llama3.1-IVON-SFT-8B

BayesRL/Llama3.1-IVON-SFT-8B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

2,4980
8.0B24 GB+ VRAMapache-2.0
Deployment details