Twu31/Qwen3.8-27B-AWQ-INT4-MTP-LowLatency
Twu31/Qwen3.8-27B-AWQ-INT4-MTP-LowLatency is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
414,302 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
Twu31/Qwen3.8-27B-AWQ-INT4-MTP-LowLatency is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
protectai/distilroberta-base-rejection-v1 is a text classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
abhishekchohan/Qwen3.8-27B-GPTQ-INT4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
prism-ml/Ternary-Bonsai-27B-AWQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
weareapexcreators/LFM2.5-1.2B-Instruct-LocalAI is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.8-27B-3MPER0RR-abliterated-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
QuaduxIT/Qwen3.8-27B-Whitehat-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
nurdich/Qwen3.8-27B-Uncensored-HauhauCS-Aggressive-MTP-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
wangzhang/Qwen3.6-35B-A3B-abliterated-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
mradermacher/Huihui-ThinkingCap-Qwen3.6-27B-abliterated-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
SpeakoFlow/speakoflow-mini is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
qaqab/Qwen3.8-27B-Uncensored-Q4_K_M-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
dahara1/gemma-4-12B-it-qat-UD-japanese-imatrix is a any to any model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
Ostfralla/Qwen3.8-27B-NVFP4-NInfer is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
weareapexcreators/LFM2.5-230M-LocalAI is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
yachen4ever/Qwen3.8-4B-Distill-Heretic-Abliterated-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
voyageai/voyage-context-4 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
gokaygokay/Flux-Prompt-Enhance is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
SC117/Ling-3.0-flash-abliterated-APEX-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Luis23333/Qwen3.8-27B-SSMFIX-UD-Q3_K_XL-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
sallani/EUAIAct-Qwen2.5-0.5B-Edge is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
rapid-mlx/Ling-3.0-tiny-MLX-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
unsloth/gemma-4-12b is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
mradermacher/Iris-12B-v1.4.1-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.