Qwen/Qwen2.5-Coder-32B-Instruct
Qwen/Qwen2.5-Coder-32B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
200,823 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
Qwen/Qwen2.5-Coder-32B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
Comfy-Org/Qwen-Image-Edit_ComfyUI is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
deepseek-ai/DeepSeek-V3.2 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Qwen/Qwen3-8B-AWQ is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
google/flan-t5-base is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
farbodtavakkoli/OTel-LLM-E4B-IT is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mudler/ced-gguf is a audio classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
google/siglip2-base-patch16-224 is a zero shot image classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
HauhauCS/Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
rhasspy/faster-whisper-tiny-int8 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Qwen/Qwen3.5-35B-A3B-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
Qwen/Qwen2.5-Coder-32B-Instruct-AWQ is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
ornith-ai/Ornith-1.0-9B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
prajjwal1/bert-tiny is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
answerdotai/ModernBERT-large is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
QuantTrio/Qwen3-VL-30B-A3B-Instruct-AWQ is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
distilbert/distilroberta-base is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
jonatasgrosman/wav2vec2-large-xlsr-53-finnish is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
microsoft/deberta-v3-large is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
jonatasgrosman/wav2vec2-large-xlsr-53-chinese-zh-cn is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
HuggingFaceTB/SmolVLM2-500M-Video-Instruct is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mistralai/Mistral-7B-Instruct-v0.2 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
cambridgeltl/SapBERT-from-PubMedBERT-fulltext is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
NbAiLab/nb-wav2vec2-1b-nynorsk is a automatic speech recognition model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.