williamliao/gemma-4-12B-it-DFlash-GGUF
williamliao/gemma-4-12B-it-DFlash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
993,047 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
williamliao/gemma-4-12B-it-DFlash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
cstr/qwen3-tts-1.7b-customvoice-GGUF is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
deucebucket/KAT-Coder-V2.5-Dev-Cerebellum-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3-4B-Instruct-2507-uncensored-v2-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
mradermacher/MS3.2-PaintedFantasy-v4.1-24B-ultra-uncensored-heretic-v2-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mobilint/Qwen3-VL-2B-Instruct is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Audio8/audio8-TTS-0.1B-ONNX-INT8 is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mradermacher/Human-Like-Mistral-Nemo-Instruct-2407-MPOA-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
prism-ml/Ternary-Bonsai-27B-unpacked is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
cvgro/Qwen3.8-27B-ABLITERATED-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Alberto-Codes/gemma-4-31B-it-fit24gib-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3-8B-DAG-Reasoning-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.5-9B-Claude-4.6-HighIQ-INSTRUCT-HERETIC-UNCENSORED-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mondk/GGUF.chatgpt-gpt-codex-V2 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
ASLP-lab/DiffRhythm2 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
PassingByPixels/Qwen3.8-27B-heretic-ara-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
kernels-community/flash-attn4 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
alphaedge-ai/multilingual-e5-large-nep-16384 is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3-VL-8B-V6-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Hebrew_Nemo-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/qwen36-35B-opus-reasoning-merged-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
glacier-hf/GLACIER-100k-MiniMol is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mradermacher/Ornith-1.5-35B-A3B-heretic-ja-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
guell00/Nexora-Gemma-4-Coder is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.