bartowski/TheDrummer_Behemoth-128B-v3-GGUF
bartowski/TheDrummer_Behemoth-128B-v3-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 96 GB. It is publicly listed on Hugging Face.
410,302 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
bartowski/TheDrummer_Behemoth-128B-v3-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 96 GB. It is publicly listed on Hugging Face.
PiehSoft/Qwen3.6-40B-Deckard-MTP is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
offmonreal/Ornith-1.5-35B-MaxQuality-MTP-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
bartowski/Qwen_Qwen3.5-397B-A17B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.
AutomatosX/AX-Qwen3.8-2.4T-A95B-MLX-AXQ-2bit-MTP is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
rizzoaiacademy/rizzo-pii-0.3B is a token classification model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
reaperdoesntknow/Qwen3.5-2B-Opus-Distil is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
prithivMLmods/Qwen3-VL-8B-Instruct-c_abliterated-v3 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. Access approval is required on Hugging Face.
openbmb/BitCPM-CANN-8B-unquantized is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/qwen-3.8-27b-abliterated-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
aldenw/Qwen3.8-27B-Uncensored-Aggressive-i1-IQ4_XS-Smaller-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
open-gigaai/Giga-World-Policy-0.5 is a robotics model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
AxionML/Gemma-4-12B-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
mradermacher/Muse-Glimmer-30B-Hermes-Agentic-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Brian6145/Qwen3.6-27B-Claude-Opus-DeepSeek-Distilled-Imatrix-MTP-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Foresee/Qwen3.8-9B-heretic-uncensored-5bit-MLX is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/For-Her-Darkside-12B-v1.4-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
tepirale/Ornith-Agents-A1-3.7-35B-A3B-dare_ties_v4-MTP-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
robbyant/lingbot-world-fast-diffusers is a image to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Hypnos-i1-8B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
TelperionAI/Huihui-Qwen3.8-27B-abliterated-INT4-AWQ-GPTQ is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
lemuralabs/Qwen3.6-27B-Claude-Opus-Reasoning-Distill-v2-abliterated-OptiQ-3.7bpw-mlx is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
lyf/Qwen3.8-27B-Blackfrost-Abliterated-NVFP4-MTP-VL is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
Adnan666/whisper-small-pashto-run11-cv24only is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.