unsloth/Qwen3.6-27B-UD-MLX-NVFP4
unsloth/Qwen3.6-27B-UD-MLX-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
1,035,045 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
unsloth/Qwen3.6-27B-UD-MLX-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
ai-sage/GigaChat3.1-10B-A1.8B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Shakker-Labs/AWPortrait-Z is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/medgemma-27b-it-i1-GGUF is a image classification model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
leonsarmiento/Qwen3.8-27B-3bit-mtp-mlx is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
unsloth/ERNIE-Image-GGUF is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mlx-community/Qwen-Image-2512-6bit is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
JohnTdi/Mage-VL-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
DoorZekor/WAN2.2-14B-Rapid-AllInOne-GGUF-NSFW-v10 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
bartowski/allenai_Olmo-3-7B-Think-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
seanbailey518/Huihui-GLM-4.6V-Flash-abliterated-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3-0.9B-A0.6B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
nvidia/dvlt is a image to 3d model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3-Reranker-4B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
inclusionAI/Ling-3.0-tiny-base is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
KeefeBuild/Keefe-Discere is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
vibe-app/diar-streaming-sortformer-4spk-v2-gguf is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
zeromodels/gemma-4-e4b-it is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
alexokita/Darkhn-M3.2-36B-Animus-V12.0-Heretic-Uncensored-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
mradermacher/Mistral-Small-24B-Instruct-Jbliterated-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Tele-AI/TeleChat3-36B-Thinking is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
mlx-community/gemma-4-e4b-it-qat-OptiQ-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
WPS-Qingqiu/OmniAlign is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.5-9B-Unredacted-MAX-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.