unsloth/gemma-4-E2B-it-UD-MLX-4bit
unsloth/gemma-4-E2B-it-UD-MLX-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
651,070 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
unsloth/gemma-4-E2B-it-UD-MLX-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/RIFA-Edge-0.6B-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3-30B-A3B-Instruct-2507-Heretic-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
cstr/qwen3-asr-1.7b-ja-anime-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
Wespeaker/wespeaker-voxceleb-redimnet2-B6-LM is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
JC1DA/Qwen3.8-27B-heretic-ara-W4A16 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/InSight-doc-8B-i1-GGUF is a reinforcement learning model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.5-122B-A10B-abliterated-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 96 GB. It is publicly listed on Hugging Face.
unsloth/GLM-5.1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.8-27B-absolute-heresy-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
reaperdoesntknow/Qwen3.5-2B-CyberSec is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/gpt-oss-20b-Uncensored-xCloud-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 16 GB. It is publicly listed on Hugging Face.
mradermacher/ExoMind-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
cstr/vibevoice-asr-bitnet-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
AxiomicLabs/GPT-X2.5-135M is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Darwin-9B-NEG-x-Negentropy-V8-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.5-13B-Deckard-Heretic-Uncensored-Thinking-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
DavidAU/Qwen3.6-27B-NEO-CODE-Di-IMatrix-MAX-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
lyf/Qwen3.8-27B-Huihui-Abliterated-NInfer-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3-1.6B-A0.9B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
PassingByPixels/Qwen3.6-27B-Architect-Polaris2-Fable-B-F451-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
RedHatAI/Qwen2.5-7B-Instruct-FP8-dynamic is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Mungert/Qwen3.8-27B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Guile/orcarouter_Qwen3.8-27B-Uncensored-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.