mlx-community/GLM-4.7-6bit
mlx-community/GLM-4.7-6bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
1,546,995 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
mlx-community/GLM-4.7-6bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/TAMA-vA-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/zen4-coder-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.5-4B_Abliterated-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
mradermacher/gpt-oss-20b-heretic-ara-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 16 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.5-4B-Claude-Opus-Reasoning-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.5-0.8B-uncensored-ara-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
bekoozkan/godot-qwen2.5-coder-7b-instruct-bnb-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mlx-community/Qwen3.6-27B-MTP-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.8-27B-Opus-Abliterix-Reasoning-BF16-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/T3Q-qwen2.5-14b-v1.0-e3-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
wangkanai/qwen3-vl-8b-instruct is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/salamandra-7b-instruct-tools-20260108-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3-30B-A3B-Thinking-2507-GLM-4.7-Flash-High-Reasoning-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/TimeOmni-1-7B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
MIRALABS/Ornith-1.5-35B-A3B-W4A16-SYM is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
TaoLiveAIGC/TLive-Omni-9B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
pragmaticcs/SignOfFour-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/D2IL-Japanese-Qwen2.5-32B-Instruct-v0.1-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
CodeGoat24/UnifiedReward-2.0-qwen3vl-32b is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
mradermacher/Math-Coma-7B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Diffusion-Llama-3-8B-Instruct-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Xenova/sam-vit-base is a mask generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mpasila/Qwen3.5-4B-EU-Q4_K_M-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.