avlp12/GLM-5.3-Flash-Alis-MLX-4bit
avlp12/GLM-5.3-Flash-Alis-MLX-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
1,624,985 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
avlp12/GLM-5.3-Flash-Alis-MLX-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3-Coder-30B-A3B-Instruct-480B-Distill-V2-Fp32-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/gemma-4-E4B-it-uncensored-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
MoonRide/gemma-4-26B-A4B-it-heretic-ara-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
wangyue114514/rwkv7-g1d-0.1b-hf is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
reaperdoesntknow/DualMind-TKD-Agentic-1.7B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mlc-ai/Qwen2.5-0.5B-Instruct-q4f32_1-MLC is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mradermacher/QiMing-PR-20B-MXFP4-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 16 GB. It is publicly listed on Hugging Face.
mradermacher/CrucibleLab-L3.3-70B-Loki-V2.0-Heretic-Uncensored-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
huihui-ai/Huihui-Ornith-1.5-9B-abliterated is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/Dirty-Shirley-Writer-v0-abliterated-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/gemma-3-12b-it-heretic-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/gpt-oss-20b-absolute-heresy-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 16 GB. It is publicly listed on Hugging Face.
meshllm/gemma-4-31B-it-qat-UD-Q4_K_XL-layers is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
vcruz305/Ornith-1.5-35B-A3B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
Mungert/NVIDIA-Nemotron-Nano-12B-v2-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/DukunLM-13B-V1.0-Uncensored-sharded-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
machiabeli/Qwen-Image-2512-4bit-MLX is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
tiny-random/glm-4.7-flash is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
bartowski/stepfun-ai_Step-3.5-Flash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
geoffmunn/Qwen3-32B-f16 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Nexuss0781/Ethio-BBPE is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
lightseekorg/kimi-k2.5-eagle3-mla is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
gokaygokay/Florence-2-Flux-Large is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.