amd/gpt-oss-20b-MoE-Quant-W-MXFP4-A-FP8-KV-FP8
amd/gpt-oss-20b-MoE-Quant-W-MXFP4-A-FP8-KV-FP8 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
452,288 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
amd/gpt-oss-20b-MoE-Quant-W-MXFP4-A-FP8-KV-FP8 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
Rabbit-Hole-Ai/Qwen3.8-27B-5090-goldilocks-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
bottlecapai/ThinkingCap-Qwen3.6-27B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. Access approval is required on Hugging Face.
cusiman/Krea2-realism-V2 is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
bartowski/ProCreations_grug-35b-v2-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
rzgar/minimax_h3_ref2va_fp8_e4m3fn is a image to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
PeppX/gemma-4-e2b-uncensored-litertlm is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
reaperdoesntknow/Dualmind-Qwen-1.7B-Thinking is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
FunAudioLLM/Fun-ASR-Nano-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
furiosa-ai/Qwen3-Reranker-4B is a text classification model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/Mistral-NeMo-12B-Abliterated-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
cross-encoder/ettin-reranker-1b-v1 is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.8-27B-Uncensored-Cyber-i1-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
bartowski/bottlecapai_ThinkingCap-Qwen3.6-27B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
sanrenti/Flux2-Klein-9B-True-V3 is a text to image model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
orcarouter/Qwen3.8-27B-Uncensored-INT8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. Access approval is required on Hugging Face.
unsloth/Qwen3.5-397B-A17B-MTP-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.
bartowski/migtissera_Tess-4-27B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/Huihui-Qwen3.6-27B-abliterated-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Atomic-Germ/Qwen3.8-Distilled-1.2B-NPU2 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mradermacher/SearchQwen2.5-7B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/4BeastsOfApocalypse-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
bartowski/ATH-MaaS_OvisOCR2-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
vcruz305/GLM-5.3-Flash-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.