baicai1145/DeepSeek-V4-Flash-0731-W4A16
baicai1145/DeepSeek-V4-Flash-0731-W4A16 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
663,069 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
baicai1145/DeepSeek-V4-Flash-0731-W4A16 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Akahsizrr/fuse-1-Lite-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
inclusionAI/Ling-lite-1.5 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Brian6145/Qwen3.6-27B-Claude-Opus-DeepSeek-Distilled-Imatrix-MTP-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Foresee/Qwen3.8-9B-heretic-uncensored-5bit-MLX is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
jamesrogers/Qwen3.8-Flash-Next-MTP-MXFP4-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/For-Her-Darkside-12B-v1.4-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
TestOrganizationPleaseIgnore/WAMU_v3_WAN2.2_I2V_LIGHTNING is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
tepirale/Ornith-Agents-A1-3.7-35B-A3B-dare_ties_v4-MTP-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
inclusionAI/Ming-flash-omni-2.0 is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.8-27B-Kimiko-2-BF16-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
tiny-random/gemma-4-dense is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
InstaDeepAI/NTv3_650M_post is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.
robbyant/lingbot-world-fast-diffusers is a image to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Hypnos-i1-8B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
TelperionAI/Huihui-Qwen3.8-27B-abliterated-INT4-AWQ-GPTQ is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
unsloth/NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.
filvyb/Qwen3.5-9B-heretic-v2-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mainguyen9/vietlegal-harrier-0.6b is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
lemuralabs/Qwen3.6-27B-Claude-Opus-Reasoning-Distill-v2-abliterated-OptiQ-3.7bpw-mlx is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
lyf/Qwen3.8-27B-Blackfrost-Abliterated-NVFP4-MTP-VL is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
Adnan666/whisper-small-pashto-run11-cv24only is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
vcruz305/Qwen3.8-27B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
cyankiwi/Ornith-1.0-35B-AWQ-INT4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.