furiosa-ai/Qwen3-VL-2B-Thinking
furiosa-ai/Qwen3-VL-2B-Thinking is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
396,301 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
furiosa-ai/Qwen3-VL-2B-Thinking is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
saga404/Qwen3.8-9B-heretic-uncensored-Q5_0-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
cyankiwi/Ornith-1.5-9B-AWQ-INT4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
llm-semantic-router/multi-modal-embed-small is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
ManniX-ITA/gemma-4-A4B-98e-v7-coder-it-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
vidore/colSmol-256M is a visual document retrieval model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
kernels-community/mamba-ssm is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
orcarouter/Qwen3.8-Flash-Next-Uncensored-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.
mradermacher/KAT-Coder-V2.5-Dev-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.6-27B-Jormungandr-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
detrax/Qwen3-4B-Thinking-2507-Qwen3.8-Max-Distillation-Detrax is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
vantagewithai/Bernini-R-GGUF-ComfyUI is a image text to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
reaperdoesntknow/TopologicalQwen is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
llm-jp/llm-jp-4-33b-thinking-gguf is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
cyankiwi/Ornith-1.5-35B-A3B-AWQ-INT4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
magiccodingman/Qwen3.8-27B-MXFP4-MagicQuant-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Blasserman/Qwen3.6-9B-Heretic-Uncensored-Thinking-Sweet-Madness-Q4_K_M-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
NANI-Nithin/Mellum2-12B-A2.5B-Instruct-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/Melody1437-31B-v2.0-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
open-gigaai/GigaBrain-0.7-3.5B-Base is a robotics model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
Shiftedx/Qwen3.8-27B-Abliterated-MLX-MXFP4-MTP is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
esherialabs/saferide-gemma-4-e2b-v058-original-419806-litertlm is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
google/gemma-4-26B-A4B-it-qat-q4_0-unquantized-assistant is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Vontra/GLM-5.3-Flash-MLX-4bit-MTP is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.