coolthor/gemma-4-12B-it-FP8-dynamic
coolthor/gemma-4-12B-it-FP8-dynamic is a any to any model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
408,301 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
coolthor/gemma-4-12B-it-FP8-dynamic is a any to any model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
kai-os/Carnice-V3-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
syvai/qwen3.8-27b-3090-fast-variant is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
bartowski/nex-agi_Nex-N2-mini-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
pottokao/Ornith-1.5-35B-A3B-abliterated-NVFP4-DFlash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
apodex/Apodex-1.1-mini-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
openbmb/BitCPM-CANN-1B-unquantized is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
openbmb/BitCPM-CANN-0.5B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
intelligenAI/intellifold is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
AIOpsInSpace/Qwen2.5-Coder-14B-Instruct-Uncensored-Patched is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
reaperdoesntknow/Qwen3-1.7B-Thinking-Distil is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
jackklki990/DeepSeek-R1-Distill-Qwen-7B-Uncensored-Reasoner-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
speakleash/Bielik-11B-v3.0-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. Access approval is required on Hugging Face.
hozifa1/Faqih-R1-14B-Islamic-AI is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
donghanasd/Huihui-Qwen3.8-27B-abliterated-KO-Ridge-3.7bpw-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
xero0000/Qwen3.6-35B-A3B-128E-Pruned-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.6-21B-IQ-Ultra-Heretic-Uncensored-Thinking-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 16 GB. It is publicly listed on Hugging Face.
onnx-community/bge-m3-ONNX is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
bartowski/TheDrummer_Behemoth-128B-v3-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 96 GB. It is publicly listed on Hugging Face.
PiehSoft/Qwen3.6-40B-Deckard-MTP is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
offmonreal/Ornith-1.5-35B-MaxQuality-MTP-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
bartowski/Qwen_Qwen3.5-397B-A17B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.
AutomatosX/AX-Qwen3.8-2.4T-A95B-MLX-AXQ-2bit-MTP is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
rizzoaiacademy/rizzo-pii-0.3B is a token classification model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.