code-q3_1p7b-replay-lam0p1
RL-Forgetting-Experiments-3/code-q3_1p7b-replay-lam0p1
Reinforcement LearningLicense: apache-2.0
Tracked by OpenModelStats since October 1, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 0
- Total downloadsCumulative all-time download counter reported by the source platform.
- 0
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- —
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 0
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads.
- —
- –this week
- Parameters
- —
- Last source update
- Oct 5, 2026
Downloads over time
Daily gain between consecutive observationsLikes over time
Rank among tracked models
Lower is betterNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-10-01. Charts appear as daily observations accumulate.
Overview
code-q3_1p7b-replay-lam0p1 is a reinforcement learning model published by RL-Forgetting-Experiments-3. OpenModelStats has tracked the model since Oct 1, 2026, most recently observing it Oct 6, 2026.
- Published
- Oct 1, 2026
- Last updated
- Oct 5, 2026
- Library
- —
- License
- apache-2.0
- Task
- Reinforcement Learning
- Parameters
- —
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- —
safetensorsreinforcement-learninggrpocode-generationmbppcatastrophic-forgettingbase_model:Qwen/Qwen3-1.7B-Basebase_model:finetune:Qwen/Qwen3-1.7B-Baselicense:apache-2.0region:us
Publisher
RL-Forgetting-Experiments-3View publisher statistics →Derived from Qwen/Qwen3-1.7B-Base (reported by the source; parent not yet tracked).
Similar models
| Model | DL 30D |
|---|---|
| ppo-seals-CartPole-v0HumanCompatibleAI | 171.8K |
| ppo-Pendulum-v1HumanCompatibleAI | 46K |
| Tifa-Deepsex-14b-CoT-GGUFmradermacher | 20.9K |
| NoveltyAdapt-policieshelenlu94 | 20.7K |
| sac-BipedalWalkerHardcore-v3sb3 | 8,773 |
| Qwen3-14B-ARPO-DeepSearch-i1-GGUFmradermacher | 7,238 |