Vontra/DeepSeek-V4-Flash-0731-MXFP4-MLX
Vontra/DeepSeek-V4-Flash-0731-MXFP4-MLX is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
304,304 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
Vontra/DeepSeek-V4-Flash-0731-MXFP4-MLX is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
bartowski/sophosympatheia_Glistening-Gem-31B-v2.0-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
KristianS7/Ouro-1.4B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
Abiray/OvisOCR2-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Tencent-Hy-30B-A3B-uncensored-heretic-i1-GGUF is a translation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Youssofal/Qwen3.8-27B-MTPLX-Bare-Speed is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
dealignai/Muse-Glimmer-30B-CRACK-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
tsystems/colqwen2.5-3b-multilingual-v1.0 is a visual document retrieval model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/Juicer-35B-A3B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
llmfan46/gemma-4-12B-it-uncensored-heretic-GGUF is a any to any model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mudler/Gemopus-4-26B-A4B-it-Preview-APEX-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/JoyFox-Qwen3.6-35B-A3B-RP-Aggressive-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
unsloth/Phi-3.5-mini-instruct-bnb-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
eremeev-d/graphpfn-1.3 is a graph ml model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
BennyDaBall/Qwen3-4b-Z-Image-Engineer-V4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
coolthor/comfyui-zimage-sulphur-nvfp4 is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mudler/Qwopus3.6-35B-A3B-Coder-APEX-MTP-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
esatapedico/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NM-DAU-NVFP4-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
ocminer/Qwen3-8B-9c925d64-bf16-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
ReadyArt/Omega-Convergence-27B-v1.0-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
LiconStudio/LTX-2.5-Multiple-Subject-Reference is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
cstr/qwen3-asr-1.7b-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mlx-community/DeepSeek-V4-Flash-0731-OptiQ-2bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mudler/Mistral-Small-4-119B-2603-APEX-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.