Tradeoffs, side by side.
Select up to four models. These signals narrow the shortlist; benchmark on your own workload before deploying.
| Signal | DGurgurov/OpenThinker3-7B-SFT-GRPO-DE |
|---|---|
| Task | reinforcement-learning |
| Parameters | 7.0B |
| Minimum VRAM | 24 GB |
| Recommended VRAM | 48 GB |
| License | Not declared |
| Commercial signal | VERIFY_LICENSE |
| Downloads | 2,134 |
| Likes | 0 |
| Source | Hugging Face ↗ |