cross-encoder/ettin-reranker-400m-v1
cross-encoder/ettin-reranker-400m-v1 is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
675,866 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
cross-encoder/ettin-reranker-400m-v1 is a text ranking model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
scottlowry/Qwen3.8-27B-oQ4e-fp16-mtp is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
EasiiX/Qwen3.8-Flash-Next-MTP-Strix-Halo-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
PrunaAI/dolphin-2.9-llama3-70b-GGUF-smashed is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
mudler/Carnice-MoE-35B-A3B-APEX-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
unsloth/Qwen3.6-35B-A3B-MLX-8bit is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 48 GB. It is publicly listed on Hugging Face.
gbuzhf/Huihui-Ornith-1.0-35B-abliterated-MTP-UD-APEX-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
reaperdoesntknow/Qemma-sft is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mudler/LFM2.5-8B-A1B-APEX-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/gemma-4-26B-A4B-it-qat-q4_0-unquantized-uncensored-heretic-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/Gemma-4-31B-Loki-Scotoma-V2.0-Swa-4096-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mixedbread-ai/mxbai-edge-colbert-v0-32m is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
llmfan46/Qwen3.5-27B-Nikusui-v1-Uncensored-Heretic-Native-MTP-Preserved-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
chimingw/Qwen3.8-27B-AEON-ULTIMATE-UNCENSORED-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
apolo13x/Qwen3.5-9B-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
RedHatAI/GLM-5.2-speculator.dspark is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
cyankiwi/MiniMax-M3-AWQ-INT4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
ibm-granite/granite-4.0-h-tiny-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Stabhappy/gemma-4-31B-it-heretic-Gguf is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Sunbird/asr-whisper-51-african-languages is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.
aufklarer/DeepFilterNet3-CoreML is a audio to audio model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
MERaLiON/MERaLiON-3-3B-ASR is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
meshllm/GLM-4.7-UD-Q4_K_XL-layers is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/spoomplesmaxx-mockingbird-36B-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.