SafeSelfPlay-checkpoints
xudongwu/SafeSelfPlay-checkpoints
Tracked by OpenModelStats since September 5, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 0
- Total downloadsCumulative all-time download counter reported by the source platform.
- 0
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- 0
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 1
- +0 in 7D
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads.
- —
- –this week
- Parameters
- —
- Last source update
- Sep 26, 2026
Downloads over time
Cumulative total downloads observed by OpenModelStatsLikes over time
Rank among tracked models
Lower is betterNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-09-05. Charts appear as daily observations accumulate.
Overview
SafeSelfPlay-checkpoints is a reinforcement learning model published by xudongwu. OpenModelStats has tracked the model since Sep 5, 2026, most recently observing it Oct 5, 2026.
- Published
- Aug 18, 2026
- Last updated
- Sep 26, 2026
- Library
- peft
- License
- —
- Task
- Reinforcement Learning
- Parameters
- —
- Spaces
- 0
- Derivative models
- —
- Velocity 7D/day
- 0
peftsafetensorsreinforcement-learninglorasafetybase_model:mlabonne/Meta-Llama-3.1-8B-Instruct-abliteratedbase_model:adapter:mlabonne/Meta-Llama-3.1-8B-Instruct-abliteratedregion:us
Publisher
xudongwuView publisher statistics →Model family
Based on base-model relationships reported by the source's metadata.
Base model
Sibling models
Similar models
| Model | DL 30D |
|---|---|
| ppo-seals-CartPole-v0HumanCompatibleAI | 164.7K |
| ppo-Pendulum-v1HumanCompatibleAI | 44.6K |
| Tifa-Deepsex-14b-CoT-GGUFmradermacher | 20.8K |
| NoveltyAdapt-policieshelenlu94 | 20.7K |
| sac-BipedalWalkerHardcore-v3sb3 | 8,750 |
| decision-transformer-gym-hopper-mediumedbeeching | 6,551 |