unsloth/Phi-3.5-mini-instruct-bnb-4bit
unsloth/Phi-3.5-mini-instruct-bnb-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
306,303 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
unsloth/Phi-3.5-mini-instruct-bnb-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
eremeev-d/graphpfn-1.3 is a graph ml model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
BennyDaBall/Qwen3-4b-Z-Image-Engineer-V4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
coolthor/comfyui-zimage-sulphur-nvfp4 is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mudler/Qwopus3.6-35B-A3B-Coder-APEX-MTP-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
esatapedico/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NM-DAU-NVFP4-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
ocminer/Qwen3-8B-9c925d64-bf16-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
ReadyArt/Omega-Convergence-27B-v1.0-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
LiconStudio/LTX-2.5-Multiple-Subject-Reference is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
cstr/qwen3-asr-1.7b-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mlx-community/DeepSeek-V4-Flash-0731-OptiQ-2bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mudler/Mistral-Small-4-119B-2603-APEX-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
ProGamerGov/qwen-360-diffusion is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
flywheel-ai/automotive is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
z-lab/Qwen3.5-122B-A10B-DFlash is a text generation model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.
IAAR-Shanghai/Metis-4B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
SC117/Ornith-1.0-35B-MTP-APEX-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
q-future/Q-ReAlign-Pro-9B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
llmfan46/Qwen3.5-9B-Nikusui-v1-Uncensored-Heretic-Native-MTP-Preserved-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
microsoft/Phi-4-reasoning-vision-15B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
Foresee/Qwen3.8-9B-heretic-uncensored-4bit-MTPLX is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
cyankiwi/Ovis2.6-30B-A3B-AWQ-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
handy-computer/whisper-small.en-gguf is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
bartowski/ai9stars_G9v3-3B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.