RedHatAI/Qwen3-4B-FP8-dynamic
RedHatAI/Qwen3-4B-FP8-dynamic is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
1,153,035 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
RedHatAI/Qwen3-4B-FP8-dynamic is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
primitive-ai/Qwen3.8-27B-mixed-NVFP4-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
mradermacher/MathSmith-Hard-Problem-Synthesizer-Qwen3-8B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Barrrrry/DeepSeek-R1-W4AFP8 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
lyssquant/Inkling-Small-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/OpenGCM-v2-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Vontra/Qwen3.8-27B-oQ2 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
litert-community/Qwen3-4B-Instruct-2507 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/gpt-oss-20b-plan-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 16 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.5-9B-YOYO-Thinking-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
BAAI/BGE-VL-large is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mradermacher/Gemma-4-Queen-31B-it-uncensored-heretic-GGUF is a any to any model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
evsinlb/Qwen3.8-27B-oQ4e-mtp is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
ApolloRaines/Qwen2.5-Coder-14B-Instruct-Jbliterated is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.6-35B-A3B-StyleTune-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
BennyDaBall/Qwen3-4b-Z-Image-Engineer-V2.5 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
piotrmaciejbednarski/gliner2-polish-pii is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/MistralPrism-24B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/StormSeeker-24B-v1-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
OptimizeLLM/Qwen3-VL-30B-A3B-Thinking-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
mradermacher/Vidhaan-72B-Legal-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
cstr/wav2vec2-large-xlsr-53-english-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
DFveloper/gemma-4-26B-A4B-Claude-Opus-4.7-QAT-Q4_0-Heretic-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/ChatBerry-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.