sanbanfu/Qwen3.8-27B-ShortThink-NVFP4-GGUF
sanbanfu/Qwen3.8-27B-ShortThink-NVFP4-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
937,050 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
sanbanfu/Qwen3.8-27B-ShortThink-NVFP4-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
aj9o9/GLM-5.3-Flash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/MN-Nazgul-12B-v1-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
hjgycm/Qwen3.6-14B-A3B-VibeForged-v2-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
nerkyor/Qwen3.8-27B-Uncensored-Heretic-T10-BF16 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
NeuML/biomedbert-base-reranker is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/mox-8b-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Vishva007/Qwen3-VL-8B-Instruct-W4A16-AutoRound-AWQ is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
surogate/Qwen3.5-0.8B-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
macwiatrak/bacformer-masked-complete-genomes is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/NaNovel-9B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
avlp12/Qwen3.8-27B-Alis-MLX-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
kernels-community/paged-attention is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
unsloth/gemma-4-E4B-it-MLX-8bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
meituan-longcat/LongCat-Image-Dev is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
TaoLiveAIGC/TLive-Omni-4B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/Asita-8B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
nemozxy123/Huihui-Qwen3.5-9B-abliterated-AWQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
vs4vijay/Muse-Glimmer-30B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
CodeDevX/MultiModel-Small-229M is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.6-27B-abliterated-v2-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
aoiandroid/gemma-4-E2B-it-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Huihui-Qwen3-VL-8B-Thinking-abliterated-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
ctheodoris/Geneformer is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.