PrunaAI/dolphin-2.9-llama3-70b-GGUF-smashed
PrunaAI/dolphin-2.9-llama3-70b-GGUF-smashed is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
741,060 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
PrunaAI/dolphin-2.9-llama3-70b-GGUF-smashed is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
mudler/Carnice-MoE-35B-A3B-APEX-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
unsloth/Qwen3.6-35B-A3B-MLX-8bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
gbuzhf/Huihui-Ornith-1.0-35B-abliterated-MTP-UD-APEX-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
reaperdoesntknow/Qemma-sft is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mudler/LFM2.5-8B-A1B-APEX-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/gemma-4-26B-A4B-it-qat-q4_0-unquantized-uncensored-heretic-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/Gemma-4-31B-Loki-Scotoma-V2.0-Swa-4096-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mixedbread-ai/mxbai-edge-colbert-v0-32m is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
llmfan46/Qwen3.5-27B-Nikusui-v1-Uncensored-Heretic-Native-MTP-Preserved-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
chimingw/Qwen3.8-27B-AEON-ULTIMATE-UNCENSORED-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
apolo13x/Qwen3.5-9B-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
RedHatAI/GLM-5.2-speculator.dspark is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
cyankiwi/MiniMax-M3-AWQ-INT4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
ibm-granite/granite-4.0-h-tiny-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
oxide-lab/whisper-large-v3-turbo-GGUF is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Stabhappy/gemma-4-31B-it-heretic-Gguf is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Sunbird/asr-whisper-51-african-languages is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.
aufklarer/DeepFilterNet3-CoreML is a audio to audio model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
MERaLiON/MERaLiON-3-3B-ASR is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
meshllm/GLM-4.7-UD-Q4_K_XL-layers is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/spoomplesmaxx-mockingbird-36B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
offmonreal/Ornith-1.0-35B-MaxQuality-MTP-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
DevQuasar/deepseek-ai.DeepSeek-V3.2-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.