Ttimms/Ornith-1.5-35B-A3B-REAP-50-GGUF
Ttimms/Ornith-1.5-35B-A3B-REAP-50-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 32 GB. It is publicly listed on Hugging Face.
What to verify before deployment
- Confirm the exact weight format and quantization.
- Measure memory at your intended context or image size.
- Review model-card limitations and evaluation methodology.
- Test latency and throughput on your target runtime.
Likely commercial-friendly
License metadata is an index signal, not legal advice. Follow the repository license and any model-specific acceptable-use terms.
Verify at the source →More text-generation models
Qwen/Qwen3-0.6B
Qwen/Qwen3-0.6B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 4 GB. It is publicly listed on Hugging Face.
trl-internal-testing/tiny-Qwen2ForCausalLM-2.5
trl-internal-testing/tiny-Qwen2ForCausalLM-2.5 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
openai-community/gpt2
openai-community/gpt2 is a text generation model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
Qwen/Qwen3-8B
Qwen/Qwen3-8B is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
unsloth/Qwen3-Coder-30B-A3B-Instruct-GGUF
unsloth/Qwen3-Coder-30B-A3B-Instruct-GGUF is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Qwen/Qwen2.5-7B-Instruct
Qwen/Qwen2.5-7B-Instruct is a text generation model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Operate a GPU cloud or inference API?
Reach developers after they have selected a model and are ready to run it.