cyankiwi/gemma-4-26B-A4B-it-qat-AWQ-INT4
cyankiwi/gemma-4-26B-A4B-it-qat-AWQ-INT4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
214,822 public models indexed from the Hugging Face Hub, with independent VRAM estimates and license signals.
cyankiwi/gemma-4-26B-A4B-it-qat-AWQ-INT4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Comfy-Org/flux1-kontext-dev_ComfyUI is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Comfy-Org/stable-diffusion-3.5-fp8 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
unsloth/gemma-4-E4B-it-NVFP4 is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
DeepBeepMeep/Qwen_image is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
nvidia/Nemotron-3-Embed-8B-BF16 is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
fastino/gliner2-privacy-filter-PII-multi is a token classification model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
DeepBeepMeep/Wan2.2 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
cyankiwi/gemma-4-E4B-it-AWQ-INT4 is a any to any model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
NeuML/colbert-bert-tiny is a sentence similarity model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
incoai/Qwen3.8-27B-DFlash2-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
openthaigpt/openthaigpt1.5-7b-instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
soyrsoyr/Qwen3.8-27B-W4A16-AWQ-GPTQ is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
AtomicChat/Ornith-1.5-35B-A3B-GGUF is a image text to text model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
dummy9996/LTX-2.5-22b-ungate is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 64 GB. It is publicly listed on Hugging Face.
neuphonic/neucodec is a audio to audio model indexed for deployment research. Estimated minimum GPU memory is 8 GB. Access approval is required on Hugging Face.
jinaai/jina-embeddings-v5-omni-small is a feature extraction model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
outsourc-e/Qwen3.8-27B-Unleashed-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Comfy-Org/ERNIE-Image is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Bucoid/Qwen3.8-27B-Heretic-Ara-16GB-VRAM-IQ4-XS-MTP-GGUF is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
jinaai/jina-colbert-v2 is a AI workload model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
bartowski/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
LiquidAI/LFM2.5-2.6B-DSpark-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
Omnico/Krea2_turbo_diff_loras is a text to image model indexed for deployment research. Estimated minimum GPU memory is 12 GB. It is publicly listed on Hugging Face.