diodel/Qwen3.5-0.8B-Q4_K_M-GGUF
diodel/Qwen3.5-0.8B-Q4_K_M-GGUF is a any to any model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
What to verify before deployment
- Confirm the exact weight format and quantization.
- Measure memory at your intended context or image size.
- Review model-card limitations and evaluation methodology.
- Test latency and throughput on your target runtime.
Likely commercial-friendly
License metadata is an index signal, not legal advice. Follow the repository license and any model-specific acceptable-use terms.
Verify at the source →More any-to-any models
unsloth/Qwen2.5-Omni-3B-GGUF
unsloth/Qwen2.5-Omni-3B-GGUF is a any to any model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
diodel/Qwen3.5-2B-Q4_K_M-GGUF
diodel/Qwen3.5-2B-Q4_K_M-GGUF is a any to any model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
bartowski/MirilAI_Miril-Drone-2B-1-GGUF
bartowski/MirilAI_Miril-Drone-2B-1-GGUF is a any to any model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
diodel/Qwen3.5-0.8B-Q4_K_M-GGUF
diodel/Qwen3.5-0.8B-Q4_K_M-GGUF is a any to any model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
sj5430/Qwen2.5-Omni-3B-Q4_K_M-GGUF
sj5430/Qwen2.5-Omni-3B-Q4_K_M-GGUF is a any to any model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
rfisdhef/Qwen2.5-Omni-3B-GGUF
rfisdhef/Qwen2.5-Omni-3B-GGUF is a any to any model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
Operate a GPU cloud or inference API?
Reach developers after they have selected a model and are ready to run it.