RedHatAI/Qwen3-30B-A3B-FP8-block
RedHatAI/Qwen3-30B-A3B-FP8-block is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
1,235,024 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
RedHatAI/Qwen3-30B-A3B-FP8-block is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
MuyeHuang/DuplexOmni is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
huihui-ai/Huihui-Qwen3.5-2B-abliterated is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
litert-community/SmolLM2-135M-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Content-AI/Qwen3.5-4B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
dnotitia/DNA-VL-STEER-2B is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
lightonai/LightOnOCR-2-1B-bbox-base is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
Emiliosbs/Ben3.0-7B-Uncensored is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/ANITA-NEXT-24B-Dolphin-Mistral-UNCENSORED-ITA-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
typhoon-ai/typhoon2.1-gemma3-4b is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
andyjack/Huihui-Qwen3.6-35B-A3B-abliterated-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
tarruda/Qwen3.8-Flash-Next-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
litert-community/yolox-nano-litert is a object detection model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Gemma-4-12B-StyleTune-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
khazarai/Qwen3-4B-Kimi2.5-Reasoning-Distilled-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
sallani/PrivaMesh is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mlx-community/Qwen3.5-4B-MTP-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
abhishekchohan/Qwen3.8-27B-AWQ-INT4-FP8KV is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/Gemma-The-Writer-9B-HERETIC-Uncensored-Abliterated-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/G4-Moonlight-Dusk-26B-A4B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mlx-community/Qwen3.8-27B-oQ6 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
philbert440/Qwen3.8-27B-Uncensored-Aggressive is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
ilintar/moss-tts-gguf is a text to speech model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
QuantFactory/Qwen3-0.6B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.