SC117/gemma-4-E4B-it-heretic-QAT-GGUF
SC117/gemma-4-E4B-it-heretic-QAT-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
919,057 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
SC117/gemma-4-E4B-it-heretic-QAT-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
inclusionAI/UI-Venus-2-9B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
IQuestLab/IQuest-Coder-V1-40B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 96 GB. It is publicly listed on Hugging Face.
ibm-granite/granite-4.1-30b-fp8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Vontra/Qwen3.8-Flash-Next-MLX-8bit-MTP is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
bartowski/InternScience_Agents-A1-4B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
cstr/kokoro-voices-GGUF is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
bartowski/kai-os_Carnice-V3-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
CPSPX/babylm-zho-pinyin-code-97M is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mlx-community/Qwen3-Coder-Next-bf16 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Asmodeus-24B-v3-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Jackrong/Qwen3.5-9B-DeepSeek-V4-Flash-MTP-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3-VL-Embedding-2B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
Atomic-Germ/Ornith-1.0-9B-NPU2 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Hatim2221/Mubsir-Qwen-2B-VL is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
facebook/EUPE-ViT-S is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
gliner-community/gliner_small-v2.5 is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Janvitos/gemma-4-12B-it-qat-assistant-MTP-Q8_0-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mlx-community/KAT-Coder-V2.5-Dev-OptiQ-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
OpenMed/OpenMed-PII-ClinicalE5-Large-335M-v1 is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
frothywater/kanade-25hz-clean is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Thinking-with-Map-30B-A3B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/Ministral-3-14B-Instruct-2512-BF16-SOM-MPOA-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
byteshape/Qwen3-Coder-30B-A3B-Instruct-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.