rr-v2
vlabki/rr-v2
Tracked by OpenModelStats since September 8, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 31
- Total downloadsCumulative all-time download counter reported by the source platform.
- 31
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- 0
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 0
- +0 in 7D
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads.
- —
- –this week
- Parameters
- 898.3K
- Last source update
- Sep 5, 2026
Downloads over time
Cumulative total downloads observed by OpenModelStatsLikes over time
Rank among tracked models
Lower is betterNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-09-08. Charts appear as daily observations accumulate.
Overview
rr-v2 is a reinforcement learning model published by vlabki. OpenModelStats has tracked the model since Sep 8, 2026, most recently observing it Sep 7, 2026.
- Published
- Sep 5, 2026
- Last updated
- Sep 5, 2026
- Library
- pytorch
- License
- —
- Task
- Reinforcement Learning
- Parameters
- 898,342
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- 0
pytorchsafetensorsrr_player_recurrent_bcreinforcement-learningmario-kart-wiiregion:us
Publisher
vlabkiView publisher statistics →Similar models
| Model | DL 30D |
|---|---|
| MusaCoder-27BMooreThreads | 62 |
| jatjat-project | 185 |
| OpenThinker3-7B-SFT-GRPO-DEDGurgurov | 3,931 |
| OphVLM-R1QiZishi | 2,989 |
| Qwen3-4B-Instruct-2507.prestar-RL.reason-only.lr7e-7-kl0.step-2176JackHsieh | 334 |
| VFIG-4BXunmeiLiu | 300 |