HUGGING FACE MODEL INDEX

Search and compare Hugging Face models.

1,061,041 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.

audio-text-to-text

OpenMOSS-Team/MOSS-Audio-8B-Thinking

OpenMOSS-Team/MOSS-Audio-8B-Thinking is a audio text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 1,453♡ 83
8.0B24 GB+ VRAMapache-2.0
Deployment details →
text-generation

TrevorJS/gemma-4-12B-it-uncensored

TrevorJS/gemma-4-12B-it-uncensored is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.

↓ 1,453♡ 11
12.0B32 GB+ VRAMapache-2.0
Deployment details →
text-generation

sarvamai/sarvam-105b-fp8

sarvamai/sarvam-105b-fp8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.

↓ 1,452♡ 6
105.0B256 GB+ VRAMapache-2.0
Deployment details →
text-to-image

felipedpm/z-image-turbo-GGUF-confyui

felipedpm/z-image-turbo-GGUF-confyui is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.

↓ 1,451♡ 11
Unknown12 GB+ VRAMapache-2.0
Deployment details →
any-to-any

prithivMLmods/gemma-4-E2B-it-FP8

prithivMLmods/gemma-4-E2B-it-FP8 is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,451♡ 5
Unknown8 GB+ VRAMapache-2.0
Deployment details →
image-to-image

Alex7575/Flux2-Klein-9B-Consistency

Alex7575/Flux2-Klein-9B-Consistency is a image to image model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.

↓ 1,451♡ 1
9.0B24 GB+ VRAMapache-2.0
Deployment details →
text-generation

CortexLM/Teutonic-1-Chat-Preview

CortexLM/Teutonic-1-Chat-Preview is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,451♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text

numind/NuExtract-2.0-2B-GGUF

numind/NuExtract-2.0-2B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 1,450♡ 1
2.0B4 GB+ VRAMmit
Deployment details →
text-generation

NamanAgnih0tri/AlphaRoute-0.8B-v1.0

NamanAgnih0tri/AlphaRoute-0.8B-v1.0 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.

↓ 1,450♡ 2
800M4 GB+ VRAMapache-2.0
Deployment details →
general AI

mradermacher/Behemoth-128B-v3-GGUF

mradermacher/Behemoth-128B-v3-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 96 GB. It is publicly listed on Hugging Face.

↓ 1,449♡ 1
128.0B96 GB+ VRAMapache-2.0
Deployment details →
token-classification

DataSign/gliner-ja-pii-v1

DataSign/gliner-ja-pii-v1 is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,448♡ 2
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-generation

speleoalex/physisml-it-preview

speleoalex/physisml-it-preview is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,448♡ 0
Unknown8 GB+ VRAMmit
Deployment details →
general AI

mradermacher/AutoTriton-i1-GGUF

mradermacher/AutoTriton-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,447♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-generation

Mungert/MiroThinker-1.7-mini-GGUF

Mungert/MiroThinker-1.7-mini-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.

↓ 1,447♡ 0
Unknown8 GB+ VRAMapache-2.0
Deployment details →