VnimanieAI/Qwen3.8-Flash-Next-W4A16
VnimanieAI/Qwen3.8-Flash-Next-W4A16 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
440,300 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
VnimanieAI/Qwen3.8-Flash-Next-W4A16 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Youssofal/Qwen3.8-27B-MTPLX-Bare-Speed is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
dealignai/Muse-Glimmer-30B-CRACK-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
unsloth/NVIDIA-Nemotron-3-Ultra-550B-A55B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.
tsystems/colqwen2.5-3b-multilingual-v1.0 is a visual document retrieval model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/Juicer-35B-A3B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
llmfan46/gemma-4-12B-it-uncensored-heretic-GGUF is a any to any model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mudler/Gemopus-4-26B-A4B-it-Preview-APEX-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Applied-Innovation-Center/Karnak-40B-v1.0 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 96 GB. It is publicly listed on Hugging Face.
mradermacher/JoyFox-Qwen3.6-35B-A3B-RP-Aggressive-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
m15dg/local-ai-toolkit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
unsloth/Phi-3.5-mini-instruct-bnb-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
eremeev-d/graphpfn-1.3 is a graph ml model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
cisco-ai/SecureBERT2.0-biencoder is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
AtomicChat/Kimi-K3-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mlx-community/gemma-4-12B-it-4bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
BennyDaBall/Qwen3-4b-Z-Image-Engineer-V4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
coolthor/comfyui-zimage-sulphur-nvfp4 is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mudler/Qwopus3.6-35B-A3B-Coder-APEX-MTP-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
aisingapore/Llama-SEA-LION-v3-8B-IT-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
esatapedico/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NM-DAU-NVFP4-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
ocminer/Qwen3-8B-9c925d64-bf16-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Palmyra-Creative-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Omega-Sapphira-L3.3-70B-v1.3-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.