ddalcu/DeepSeek-V4-Flash-0731-MLX-Serve-mixed-2-3-8bit
ddalcu/DeepSeek-V4-Flash-0731-MLX-Serve-mixed-2-3-8bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
867,057 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
ddalcu/DeepSeek-V4-Flash-0731-MLX-Serve-mixed-2-3-8bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Fmuaddib/Qwen3.8-27B-AEON-ULTIMATE-UNCENSORED-BF16-mlx-8Bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
mradermacher/Huihui-Qwen3-Next-80B-A3B-Instruct-abliterated-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
singulared/DeepSeek-V4-Flash-0731-DSpark-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Jamphus/PinkCherry_NSFW_LTX23 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
CrossNow/Qwen3.8-27B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
HeartMuLa/HeartCodec-oss-20260123 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
amd/gpt-oss-20b-WFP8-AFP8-KVFP8 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
FastFlowLM/Qwen3.6-35B-A3B-NPU2 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
Codes4Fun/personaplex-7b-v1-q4_k-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
ai9stars/G9v3-3B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
devika-tiwari/gpt2_small_expandedbabyLM_175M_44 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Youssofal/Qwen3.8-Flash-Next-MTPLX-Bare-Speed is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
ACE-Step/acestep-v15-base is a text to audio model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Mistral-Nemo-2407-12B-Thinking-Claude-Gemini-GPT5.2-Uncensored-HERETIC-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
tvall43/Qwen3.5-4B-heretic-v2 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.5-27B-Unredacted-MAX-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/sarv-reasoning-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
gxcsoccer/kronos-mlx-tokenizer-base is a time series forecasting model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
turboderp/Qwen3.8-27B-exl3 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
Abiray/LTX2.3-10Eros-GGUF is a image to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/SambaLingo-Hungarian-Chat-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
LouLou1Demon/pixalium-20M-pretrained is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
EleutherAI/pythia-160m-data-seed2 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.