RukaRat/Qwen3.8-27B-INT8-W8A8-imatrix-MTP
RukaRat/Qwen3.8-27B-INT8-W8A8-imatrix-MTP is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
841,060 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
RukaRat/Qwen3.8-27B-INT8-W8A8-imatrix-MTP is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
lmstudio-community/GLM-4.7-Flash-MLX-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
INC4AI/Qwen3-8B-MXFP8-AR is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/AFM-4.5B-Base-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
mradermacher/llama3.1-heretic-unsensored-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mlx-community/LFM2.5-2.6B-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.5-9B-heretic-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
furiosa-ai/Qwen3-4B-FP8 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
laion/voiceclap-large-v2 is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
NikiKrutan/Qwen3.8-27B-MTP-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
prithivMLmods/gemma-4-E4B-it-Uncensored-MAX is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Anonym-IA/V2-camembert-ner-pii-french is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
unsloth/Krea-2-Raw is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
10Hen10/Qwen3.6-27B-unsloth-code-reasoning-deepseek-v4-pro-destilled-Mlx is a text generation model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
furiosa-ai/Qwen3-Reranker-0.6B is a text classification model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
novita/kimi-k2.7-code-eagle3-mla is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
cyankiwi/Hermes-4.3-36B-AWQ-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
bartowski/Vortex5_G4-Dark-Soul-26B-A4B-GGUF is a any to any model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/Jackdaw-3-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mlx-community/MiniMax-M2.1-4bit is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
nbpedro315/Dolphin3-Cyber-8B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
wannaphong/wav2vec2-large-xlsr-53-th-cv8-newmm is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
froggeric/Qwen3.6-27B-MTP-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
vasilpenev/nextocr-nf4-float16 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.