google/gemma-4-12B-it-qat-q4_0-unquantized-assistant
google/gemma-4-12B-it-qat-q4_0-unquantized-assistant is a any to any model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
216,821 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
google/gemma-4-12B-it-qat-q4_0-unquantized-assistant is a any to any model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
bartowski/darkc0de_Muse-Glimmer-30B-heretic-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
empero-ai/Qwen3.8-9B-Distill is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
cyankiwi/Qwen3.5-2B-AWQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
RedHatAI/diffusiongemma-26B-A4B-it-FP8-dynamic is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.8-Flash-Next-Uncensored-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
swiss-ai/Apertus-70B-Instruct-2509 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 192 GB. It is publicly listed on Hugging Face.
unsloth/gemma-4-31B-it is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
Comfy-Org/Omnigen2_ComfyUI_repackaged is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
smthem/SenseNova-U1-8B-MoT-Merger-gguf is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
codefuse-ai/F2LLM-v2-80M is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
google/gemma-4-12B-it-assistant is a any to any model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
Youssofal/Qwen3.8-27B-MTPLX-Optimized-Quality is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
RedHatAI/Qwen3.8-27B-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
philbert440/Qwen3.8-27B-Uncensored-Cyber-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
protoLabsAI/Ornith-1.5-9B-MTP-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
FINAL-Bench/Ourbox-35B-JGOS-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
AtlasCloud/DeepSeek-V4-Flash-0731-FP8-DSpark is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mudler/Qwen3.6-35B-A3B-uncensored-heretic-APEX-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
inclusionAI/Ling-3.0-tiny is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mlasli/Qwen3.8-27B-Heretic-Uncensored-Q5_K_M-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
6block/Qwen3-32B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Comfy-Org/lotus is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Comfy-Org/void-model is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.