johninthepool/Qwen3.8-27B-MTPLX-8bit
johninthepool/Qwen3.8-27B-MTPLX-8bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
1,159,035 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
johninthepool/Qwen3.8-27B-MTPLX-8bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
Myric/KAT-Coder-V2.5-Dev-APEX-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
topabaem/Qwen3.8-27B-STQ is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3-54B-A3B-2507-YOYO2-TOTAL-RECALL-Instruct-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
OpenVINO/Qwen3-8B-int4-cw-ov is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen2.5-Coder-7B-Abliterated-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
nerkyor/Qwen3.6-35B-A3B-DSV4Pro-Thinking-Distill-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
saidutta69/Qwen3-0.6B-heretic is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
specklabs/Speck1-140M-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/PocketWeights-Qwen2.5-14B-Coder-Creative-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
nvidia/Qwen3-Nemotron-235B-A22B-GenRM is a text generation model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.
onnx-community/Qwen3.5-0.8B-ONNX-OPT is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mlx-community/Qwen3.5-0.8B-OptiQ-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
logic65/whittle-next-moe-test is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Huihui-Qwen3-VL-8B-Thinking-abliterated-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Baichuan-M2-32B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/gemma-3-12b-it-vl-Kimi-V2-Heretic-Uncensored-Thinking-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.5-35B-A3B-heretic-v2-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
mradermacher/Fabled-Gemma4-31B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
majentik/Qwen3.8-27B-MLX-2bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
IvmeLabs/Ivme-Conversate-XL-v1-Base is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
devpramod-intel/granite-4.1-3b-quantized.w8a8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
mradermacher/Toucan-Qwen2.5-32B-Instruct-v0.1-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/Huihui-Ring-mini-2.0-abliterated-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.