HUGGING FACE MODEL INDEXSearch and compare Hugging Face models.
625,071 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
token-classification
EuroEval/mmBERT-small-multi-wiki-qa-synthetic-hallucinations-with-ragtruth-is is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 4,790♡ 0
Unknown8 GB+ VRAMLicense unknown
Deployment details →
image-text-to-text
Vishva007/Qwen3.8-27B-W4A16-AutoRound is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 4,786♡ 2
27.0B24 GB+ VRAMapache-2.0
Deployment details →
text-generation
mlx-community/gemma-4-12B-it-qat-OptiQ-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
↓ 4,785♡ 16
12.0B12 GB+ VRAMapache-2.0
Deployment details →
text-generation
nvidia/MiniMax-M3-DSpark is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 4,784♡ 13
Unknown8 GB+ VRAMother
Deployment details →
text-generation
furiosa-ai/Qwen3-8B-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 4,784♡ 0
8.0B24 GB+ VRAMapache-2.0
Deployment details →
general AI
dealignai/GLM-5.3-Flash-UNCENSORED-NVFP4 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 4,783♡ 28
Unknown8 GB+ VRAMmit
Deployment details →
image-text-to-text
byteshape/Qwen3.6-35B-A3B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
↓ 4,779♡ 39
35.0B32 GB+ VRAMapache-2.0
Deployment details →
text-generation
poolside/Laguna-XS-2.1-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 4,779♡ 9
Unknown8 GB+ VRAMopenmdw-1.1
Deployment details →
text-generation
mradermacher/LFM2.5-2.6B-UNCENSORED-ABLITERATED-PHILADELPHIA-CLASS-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
↓ 4,775♡ 8
2.6B4 GB+ VRAMother
Deployment details →
any-to-any
llmfan46/gemma-4-12B-it-qat-q4_0-uncensored-heretic-GGUF is a any to any model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
↓ 4,774♡ 27
12.0B12 GB+ VRAMapache-2.0
Deployment details →
automatic-speech-recognition
handy-computer/whisper-tiny-gguf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 4,774♡ 1
Unknown8 GB+ VRAMapache-2.0
Deployment details →
text-generation
RedHatAI/Llama-3.3-70B-Instruct-quantized.w4a16 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
↓ 4,774♡ 4
70.0B48 GB+ VRAMllama3.3
Deployment details →
general AI
mPLUG/GUI-Owl-1.5-8B-Instruct is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 4,771♡ 13
8.0B24 GB+ VRAMmit
Deployment details →
translation
unsloth/Hy-MT2-7B-GGUF is a translation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 4,771♡ 21
7.0B8 GB+ VRAMapache-2.0
Deployment details →
automatic-speech-recognition
vrfai/Qwen3-ASR-1.7B-fp8 is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
↓ 4,771♡ 6
1.7B6 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text
tanhuajie2001/Robo-Dopamine-GRM-2.0-4B-Preview is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
↓ 4,768♡ 1
4.0B12 GB+ VRAMapache-2.0
Deployment details →
text-to-image
QiE2035/flux2-klein-9b-uncensored-text-encoder-Q2_K-GGUF is a text to image model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 4,763♡ 3
9.0B8 GB+ VRAMother
Deployment details →
image-text-to-text
nvidia/Ising-Calibration-1.5-31B-BF16 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
↓ 4,763♡ 5
31.0B80 GB+ VRAMother
Deployment details →
text-generation
Indexnusrefather/Nyx-RP-9B-Instruct-2608-v1 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 4,762♡ 3
9.0B8 GB+ VRAMapache-2.0
Deployment details →
image-text-to-text
ulkaa/Qwen3.8-27B-AWQ-INT4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 4,761♡ 2
27.0B24 GB+ VRAMapache-2.0
Deployment details →
text-generation
lowbitcoffee/GLM-5.2-W4A16 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 4,760♡ 4
Unknown8 GB+ VRAMmit
Deployment details →
general AI
PrunaAI/Llama-3-8B-Instruct-Gradient-1048k-GGUF-smashed is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 4,758♡ 32
8.0B8 GB+ VRAMLicense unknown
Deployment details →
general AI
xFutureTechx/2026_Loras is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
↓ 4,754♡ 1
Unknown8 GB+ VRAMcreativeml-openrail-m
Deployment details →
general AI
mradermacher/Qwen3.8-27B-AEON-ULTIMATE-UNCENSORED-BF16-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
↓ 4,754♡ 2
27.0B24 GB+ VRAMapache-2.0
Deployment details →