amd/Llama-3.1-8B-Instruct-w-int8-a-int8-sym-test
amd/Llama-3.1-8B-Instruct-w-int8-a-int8-sym-test is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
1,436,996 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
amd/Llama-3.1-8B-Instruct-w-int8-a-int8-sym-test is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.8-27B-Palimpsest-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
EleutherAI/pythia-14m-seed1 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
EleutherAI/pythia-410m-seed6 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Alibaba-DAMO-Academy/RynnBrain1.1-2B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
RedHatAI/gemma-3-4b-it-quantized.w8a8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
SimpleStories/SimpleStories-V2-11M is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
sdadas/mmlw-e5-small is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
LiquidAI/LFM2.5-Audio-1.5B is a audio to audio model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
mradermacher/gemma-4-21b-a4b-it-REAP-heretic-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 16 GB. It is publicly listed on Hugging Face.
adsabs/astroBERT is a fill mask model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Huihui-gemma-4-12B-agentic-fable5-abliterated-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
unsloth/medgemma-4b-it is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
Mungert/Harbinger-24B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
r0b0tlab/Ling-3.0-flash-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
gabriellarson/Devstral-Small-2507-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
LuckyLiGY/MagicTryOn-1.3B is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 6 GB. It is publicly listed on Hugging Face.
lm-provers/QED-Nano is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
4ntoine/LocoOperator-4B-LiteRTLM is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
tardellirs/gemma3moe-16x4-270m-starter is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
timodonnell/marinfold-contacts-v1-exp199-1_5b-step145199 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 16 GB. It is publicly listed on Hugging Face.
KRLabsOrg/lettucedect-v2-mmbert-base is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
RedHatAI/Meta-Llama-3.1-405B-Instruct-FP8-dynamic is a text generation model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.
befox/Wan2.2-Animate-14B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.