local-inference-lab/GLM-5.3-Flash-NVFP4
local-inference-lab/GLM-5.3-Flash-NVFP4 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
825,860 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
local-inference-lab/GLM-5.3-Flash-NVFP4 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
ReadyArt/Serenity-12B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/Melody1437-26B-A4B-v0.4-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Jundot/Qwen3.6-35B-A3B-oQ4e-mtp is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
rafw007/qwen3-coder-next-80b-redteam-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
centml/Qwen3.6-27B-NVFP4-W4A4-mlpinf is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.8-27B-thinkingcap-abliterated-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
HoppouAI/Breeze-TTS-2.cpp is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
facebook/sam-audio-small is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.
ifmylove2011/girlslike-qweni is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
openbmb/MiniCPM-V-4.6-Thinking-gguf is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
meshllm/Qwen3-0.6B-Q4_K_M-layers is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
XHToken/Spark-X2.5-4B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
munekazu/Huihui-Qwen3.8-27B-abliterated-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
tensorblock/DeepSeek-Coder-V2-Lite-Instruct-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Runware/Pony_Diffusion_V6_XL is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
OpenVoiceOS/qwen3-forced-aligner-0.6b-q4-k-m is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
cyberdelia/CyberRealisticSemiRealPony is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
Roboflow/rf-detr-nano is a object detection model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
ReadyArt/Melody1437-12B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
huihui-ai/Huihui-gemma-4-E4B-it-abliterated is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
InternScience/Agents-A1-4B-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
nvidia/asset-harvester is a image to 3d model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Earlychildhoodeducation/EleMo-V2-Base is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.