QuantTrio/MiniMax-M2.7-AWQ
QuantTrio/MiniMax-M2.7-AWQ is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
523,674 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
QuantTrio/MiniMax-M2.7-AWQ is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/gemma-4-31b-it-3MPER0RR-abliterated-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
jingedawang/babylm-zh-run10a is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
SC117/Qwen-AgentWorld-35B-A3B-MTP-Uncensored-APEX-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
SaeedLab/MolDeBERTa-base-123M-mtr is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
llmfan46/gemma-4-26B-A4B-it-qat-q4_0-uncensored-heretic-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
LiquidAI/LFM2-350M-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
S4MPL3BI4S/gemma4-coding-agent is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
tencent/Hy-MT2-30B-A3B-FP8 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
analogalok/Qwen3.8-27B-DFlash2-Q2_K-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
CMSManhattan/JiRackUltra_1b is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
RedHatAI/Mistral-Nemo-Instruct-2407-quantized.w4a16 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
ashen-sensored/wd-eva02-tagger-2026-canary is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
OpenASR/qwen3-asr-0.6b is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
DavidAU/L3-DARKEST-PLANET-16.5B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 16 GB. It is publicly listed on Hugging Face.
Richasy/IndexTTS-2.5-GGUF is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mlasli/Qwen3.8-27B-Heretic-Uncensored-IQ4_XS-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
cyankiwi/Qwen3.6-27B-AWQ-BF16-INT8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
Koopah/Qwen3.6-35B-A3B-NVFP4 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
Shuanghai/Z-Image-Turbo_fp32-fp16-bf16_comfyui is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
LiquidAI/LFM2.5-2.6B-ONNX is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
exusiaiw/chinese-babylm-2026-v3 is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/rp-glm-4.7-flash-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
logic65/Qwen3.8-Whittle-tri-14.7B-chat is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.