grpo-summarization-reward-ablation
YuvrajSingh9886/grpo-summarization-reward-ablation
Tracked by OpenModelStats since September 17, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 0
- Total downloadsCumulative all-time download counter reported by the source platform.
- 0
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- —
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 0
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads.
- —
- –this week
- Parameters
- —
- Last source update
- Sep 17, 2026
Downloads over time
Cumulative total downloads observed by OpenModelStatsLikes over time
Rank among tracked models
Lower is betterNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-09-17. Charts appear as daily observations accumulate.
Overview
grpo-summarization-reward-ablation is a summarization model published by YuvrajSingh9886. OpenModelStats has tracked the model since Sep 17, 2026, most recently observing it Sep 18, 2026.
- Published
- May 14, 2026
- Last updated
- Sep 17, 2026
- Library
- mlx
- License
- apache-2.0
- Task
- Summarization
- Parameters
- —
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- —
mlxgrposummarizationreinforcement-learningreward-ablationenlicense:apache-2.0region:us
Similar models
| Model | DL 30D |
|---|---|
| bart-large-cnnfacebook | 1.4M |
| distilbart-cnn-12-6sshleifer | 441.1K |
| pegasus-xsumgoogle | 164.4K |
| bart-large-cnn-samsumphilschmid | 78.3K |
| long-t5-tglobal-base-sci-simplifypszemraj | 50.4K |
| MEETING_SUMMARYknkarthick | 42.9K |