beaver-7b-v1.0-reward
PKU-Alignment/beaver-7b-v1.0-reward
Tracked by OpenModelStats since September 5, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 2,737
- Total downloadsCumulative all-time download counter reported by the source platform.
- 103.3K
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- +573(increase)
- 0.6%(up) vs baseline
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 17
- +0 in 7D
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads. #23 among tracked Reinforcement Learning models.
- #17,051
- 629(down 629 places)this week
- Parameters
- 6.6B
- Last source update
- Apr 20, 2024
- MomentumOpenModelStats' composite rising-model signal (0–100): a percentile blend of absolute 7-day download gains, capped percentage growth, and likes gained. See Methodology.
- 47.0
- beaver-7b-v1.0-reward gained 573 downloads during the last seven tracked days.
- It moved from #18 to #23 among tracked Reinforcement Learning models this week.
Downloads over time
Cumulative total downloads observed by OpenModelStatsLikes over time
Rank among tracked models
Lower is betterOverview
beaver-7b-v1.0-reward is a reinforcement learning model published by PKU-Alignment. OpenModelStats has tracked the model since Sep 5, 2026, most recently observing it Oct 4, 2026.
- Published
- Jul 8, 2023
- Last updated
- Apr 20, 2024
- Library
- safe-rlhf
- License
- —
- Task
- Reinforcement Learning
- Parameters
- 6,607,351,812
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- 82
safe-rlhfsafetensorsllamareinforcement-learning-from-human-feedbackreinforcement-learningbeaversafetyai-safetydeepspeedrlhfalpacaen
Similar models
| Model | DL 30D |
|---|---|
| beaver-7b-v1.0-costPKU-Alignment | 3,278 |
| inf-retriever-v1-proinfly | 2,350 |
| Open-Reasoner-Zero-7BOpen-Reasoner-Zero | 1,811 |
| a3-rl-laion_nemotron-gym-agent-calendar-80-8Blaion | 3,383 |
| a3-rl-DCAgent_exp_rpt_e2egit-v2-10-8Blaion | 1,688 |
| VisualQuality-R1-7BTianheWu | 6,506 |