ggml-org/DeepSeek-V4-Flash-0731-GGUF
ggml-org/DeepSeek-V4-Flash-0731-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
438,300 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
ggml-org/DeepSeek-V4-Flash-0731-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
inception42/Jais-2-8B-Chat is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. Access approval is required on Hugging Face.
kernels-community/finegrained-fp8 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Luminous-Mirror-26B-A4B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
nvidia/Mistral-Medium-3.5-128B-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
LIF1014/ptdbench-Llama-3.2-1B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
LiAutoAD/Ristretto-3B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
cyankiwi/Muse-Glimmer-30B-AWQ-INT4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/GLM-4.7-Flash-REAP-39-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
gkraker04/Nanbeige4.2-3B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
sakamakismile/Huihui-gemma-4-26B-A4B-it-qat-abliterated-pad768-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
realrebelai/Krea-R-Turbo is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
Vontra/DeepSeek-V4-Flash-0731-MXFP4-MLX is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
bartowski/sophosympatheia_Glistening-Gem-31B-v2.0-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
KristianS7/Ouro-1.4B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
alexdimmock/wav2vec2-basque-100h is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Abiray/OvisOCR2-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Tencent-Hy-30B-A3B-uncensored-heretic-i1-GGUF is a translation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
joeygambino/LTX-2.5-Quantized is a text to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
VnimanieAI/Qwen3.8-Flash-Next-W4A16 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Youssofal/Qwen3.8-27B-MTPLX-Bare-Speed is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
dealignai/Muse-Glimmer-30B-CRACK-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
unsloth/NVIDIA-Nemotron-3-Ultra-550B-A55B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.