rapid-mlx/Ling-3.0-tiny-MLX-4bit
rapid-mlx/Ling-3.0-tiny-MLX-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
673,866 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
rapid-mlx/Ling-3.0-tiny-MLX-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
OpenMed/OpenMed-PII-Italian-ClinicDischarge-Base-110M-v1 is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
arcee-ai/Trinity-Large-Thinking is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Infatoshi/Qwen3.6-35B-A3B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
z-lab/gpt-oss-120b-DFlash is a text generation model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.
xFutureTechx/2024_Backups is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/Mew1-2.6B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mlboydaisuke/Nemotron-3.5-ASR-Streaming-CoreAI is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
unsloth/gemma-4-12b is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
bartowski/agentscope-ai_CoPaw-Flash-9B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Iris-12B-v1.4.1-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mlboydaisuke/Qwen2.5-Omni-3B-Audio-CoreAI is a audio text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
lmstudio-community/Qwen3-0.6B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
reaperdoesntknow/Qwen3-1.7B-Coder-Distilled-SFT is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
keithnull/Qwen3.6-35B-A3B-REAM-192-heretic-APEX-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
pearl-ai/Qwen3-30B-A3B-Instruct-2507-pearl is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
sallani/PrivaMesh-Edge is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
openbmb/BitCPM-CANN-3B-unquantized is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/Phoenix-X-26B-A4B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
vijinvinod/gemma-4-12B-agentic-fable5-composer2.5-v2-3.5x-tau2-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/gemma-4-E2B-it-heretic-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
TitanML/NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.
cross-encoder/ettin-reranker-400m-v1 is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
scottlowry/Qwen3.8-27B-oQ4e-fp16-mtp is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.