rzgar/minimax_h3_ref2va_fp8_e4m3fn
rzgar/minimax_h3_ref2va_fp8_e4m3fn is a image to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
847,060 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
rzgar/minimax_h3_ref2va_fp8_e4m3fn is a image to video model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
bartowski/zai-org_GLM-4.5-Air-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
PeppX/gemma-4-e2b-uncensored-litertlm is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
reaperdoesntknow/Dualmind-Qwen-1.7B-Thinking is a text generation model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
Forturne/Qwen3-VL-Embedding-8B-NVFP4 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
FunAudioLLM/Fun-ASR-Nano-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
furiosa-ai/Qwen3-Reranker-4B is a text classification model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/Mistral-NeMo-12B-Abliterated-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mlx-community/Qwen3-ASR-0.6B-4bit is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
cross-encoder/ettin-reranker-1b-v1 is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
trl-internal-testing/tiny-Qwen2AudioForConditionalGeneration is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.8-27B-Uncensored-Cyber-i1-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
bartowski/bottlecapai_ThinkingCap-Qwen3.6-27B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
sanrenti/Flux2-Klein-9B-True-V3 is a text to image model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
c4tdr0ut/grok-oss-Apollyon-24B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mudler/Nemotron-3-Nano-30B-A3B-APEX-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
orcarouter/Qwen3.8-27B-Uncensored-INT8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. Access approval is required on Hugging Face.
unsloth/Qwen3.5-397B-A17B-MTP-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.
mlx-community/gemma-4-12B-it-OptiQ-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
bartowski/migtissera_Tess-4-27B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/Huihui-Qwen3.6-27B-abliterated-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Atomic-Germ/Qwen3.8-Distilled-1.2B-NPU2 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mradermacher/SearchQwen2.5-7B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/4BeastsOfApocalypse-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.