AtomicChat/gemma-4-12B-it-GGUF
AtomicChat/gemma-4-12B-it-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
394,301 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
AtomicChat/gemma-4-12B-it-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mixbits/Qwen3.8-27B-NVFP4-MTP-VL-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mudler/GLM-4.7-Flash-APEX-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mlx-community/Qwen3.8-27B-Uncensored-OptiQ-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
lued/Qwen3.8-27B-huihui-abliterated-INT8-W8A16-MTP is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
bartowski/vectionlabs_Salience-27B-R5-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Hcompany/Holo-3.1-0.8B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
tencent/Hy-MT2-1.8B-1.25Bit-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
leok7v/Qwen3.8-27B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
kernels-community/gpt-oss-triton-kernels is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Gemma-4-31B-StyleTune-heretic-ara-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
enginetown/Qwen3.8-27B-Calibrated is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/gemma-4-12B-it-uncensored-heretic-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
SC117/QwenPaw-Flash-9B-heretic-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mlx-community/Qwen3.8-27B-MTP-mxfp8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
NC-AI-consortium-VAETKI/VAETKI is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
allenai/tmax-9b is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mlx-community/Muse-Glimmer-30B-8bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
Sunbird/orpheus-3b-tts-multilingual is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
qzshch/Qwen3.8-27B-Blackfrost-Abliterated-NVFP4-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
perplexity-ai/pplx-embed-context-v1-0.6b is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
CohereLabs/command-a-plus-05-2026-w4a4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
furiosa-ai/Qwen3-VL-2B-Thinking is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
saga404/Qwen3.8-9B-heretic-uncensored-Q5_0-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.