ddalcu/Qwen3.8-Flash-Next-MLX-Serve-4bit
ddalcu/Qwen3.8-Flash-Next-MLX-Serve-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
999,047 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
ddalcu/Qwen3.8-Flash-Next-MLX-Serve-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Xenova/whisper-medium is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
taide/embeddinggemma-GTAIDE-300m-2605 is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. Access approval is required on Hugging Face.
unconst/Affine-5czsc2fc98-r252-merged is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
sartajbhuvaji/GLM-4.6-Flash-text is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
GabrScar/test is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mattmireles/kokoro-coreml is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
NILKNARFGonzo/floppyx4-nonsensicalEssential-base is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
vantagewithai/Magic-Wan-Image-V2-GGUF is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
funasr/ct-punc is a text classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
TestOrganizationPleaseIgnore/WAMU-Merge-VisualEffects_WAN2.2_I2V_LIGHTNING is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
unsloth/gemma-4-E4B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
PaddlePaddle/PP-OCRv6_medium_det_safetensors is a image to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/MiLMMT-46-12B-v1.0-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
DevQuasar/LiquidAI.LFM2-2.6B-Transcript-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
unsloth/gemma-4-E2B-it-UD-MLX-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
chartreuse-verte/prose-rewriter-4b-v1.2 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
mradermacher/RIFA-Edge-0.6B-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen2.5-VL-7B-Instruct-heretic-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3-30B-A3B-Instruct-2507-Heretic-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
ench100/bodyandface is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
cstr/qwen3-asr-1.7b-ja-anime-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
Wespeaker/wespeaker-voxceleb-redimnet2-B6-LM is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
unsloth/Ministral-3-8B-Instruct-2512-unsloth-bnb-4bit is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.