embedl/Cosmos-Reason2-2B-W4A16-Edge2-FlashHead
embedl/Cosmos-Reason2-2B-W4A16-Edge2-FlashHead is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.
829,860 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
embedl/Cosmos-Reason2-2B-W4A16-Edge2-FlashHead is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.
LifetimeMistake/Qwen3-VL-Embedding-2B-AWQ-4bit is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
llmfan46/Qwen3.6-35B-A3B-uncensored-heretic is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
mradermacher/Huihui-CyberStrike-OffSec-35B-abliterated-i1-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
bartowski/SpatialAxiom_SpatialAxiom-9B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
PaddlePaddle/PP-OCRv6_tiny_rec_onnx is a image to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
bartowski/granite-4.2-3b-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
primitive-ai/Qwen3.8-Flash-Next-mixed-NVFP4-FP8 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
AEON-7/Ornith-1.0-35B-AEON-Ultimate-Uncensored-NVFP4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
HauhauCS/GPT-OSS-20B-Uncensored-HauhauCS-Balanced is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 16 GB. It is publicly listed on Hugging Face.
litert-community/Qwen3-0.6B-int4 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
longtermrisk/Qwen3-8B-german-city-names-second-third-v2-sft-seed3-epoch3 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
RemySkye/NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
QuantTrio/Qwen3.5-397B-A17B-AWQ is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 256 GB. It is publicly listed on Hugging Face.
mradermacher/qwen3.5-9b-nsfw-captioning-v2-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Qwen3.6-35B-A3B-Claude-4.7-Opus-Reasoning-Distilled-i1-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
lerobot/xvla-base is a robotics model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/Assistant_Pepe_32B-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
arissuga/aurum-brain-ai is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
IAAR-Shanghai/Metis-9B is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
batiai/Qwen3.8-Flash-Next-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Tele-AI/TeleChat2-3B is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.
kyledam/wan_lora is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
NiuTrans/LMT-60-0.6B is a translation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.