stepfun-ai/Step-3.7-Flash-GGUF
stepfun-ai/Step-3.7-Flash-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
538,273 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
stepfun-ai/Step-3.7-Flash-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
drawais/Qwen3-Reranker-0.6B-AWQ-INT4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
reaperdoesntknow/TameForCasualLM is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/gemma-4-31b-it-heretic-ara-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
ACE-Step/acestep-v15-xl-turbo-diffusers is a text to audio model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
LiquidAI/LFM2.5-Embedding-350M-GGUF is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
selode-ai/Qwen-3.6-35B-A3B-VRAP-4-bit-AWQ-21.2GB is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
z-lab/Muse-Glimmer-30B-DFlash2-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.6-27B-ABLITERATED-UNCENSORED-PHILADELPHIA-CLASS-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/L3.1-Bluesv1-8B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
rostlabs/rost-1b-instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
VoidWalkercero/Qwen3-0.6B-Particle-SousVide-R128-Perfect is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3-VL-Embedding-8B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
cyankiwi/Qwen3-Omni-30B-A3B-Instruct-AWQ-8bit is a any to any model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
mradermacher/DeepSeek-V4-Pro-Qwen3.5-9B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
QQZ2026/Qwen3.8-27B-NVFP4-Q5K-no-MTP-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
ARC4NUM/Qwen3.8-Flash-Next-Uncensored-MLX-Serve-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
akshan-main/tiny-ernie-image-modular-pipe is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Comfy-Org/Chroma1-Radiance_Repackaged is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
laion/a3-rl-laion_nemotron-gym-agent-calendar-80-8B is a reinforcement learning model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
RadixArk/Muse-Glimmer-q4k-dynamic-MLX is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
nvidia/NV-Raw2insights-MRI is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
unsloth/Qwen3.5-27B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
neuphonic/neutts-air is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.