mradermacher/KernelBench-RLVR-120b-i1-GGUF
mradermacher/KernelBench-RLVR-120b-i1-GGUF is a reinforcement learning model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
What to verify before deployment
- Confirm the exact weight format and quantization.
- Measure memory at your intended context or image size.
- Review model-card limitations and evaluation methodology.
- Test latency and throughput on your target runtime.
Likely commercial-friendly
License metadata is an index signal, not legal advice. Follow the repository license and any model-specific acceptable-use terms.
Verify at the source →More reinforcement-learning models
jamesheald/joint-space-empowerment
jamesheald/joint-space-empowerment is a reinforcement learning model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
KaptainKris/ppo-LunarLander-v3-flip
KaptainKris/ppo-LunarLander-v3-flip is a reinforcement learning model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
mradermacher/KernelBench-RLVR-120b-i1-GGUF
mradermacher/KernelBench-RLVR-120b-i1-GGUF is a reinforcement learning model indexed for deployment research. Estimated minimum GPU memory is 80 GB. It is publicly listed on Hugging Face.
laion/a3-rl-laion_nemotron-gym-agent-calendar-80-8B
laion/a3-rl-laion_nemotron-gym-agent-calendar-80-8B is a reinforcement learning model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
mradermacher/InSight-doc-8B-i1-GGUF
mradermacher/InSight-doc-8B-i1-GGUF is a reinforcement learning model indexed for deployment research. Estimated minimum GPU memory is 8 GB. It is publicly listed on Hugging Face.
DGurgurov/OpenThinker3-7B-SFT-GRPO-DE
DGurgurov/OpenThinker3-7B-SFT-GRPO-DE is a reinforcement learning model indexed for deployment research. Estimated minimum GPU memory is 24 GB. It is publicly listed on Hugging Face.
Operate a GPU cloud or inference API?
Reach developers after they have selected a model and are ready to run it.