code-llama3p2_3b-replay-lam0p1

RL-Forgetting-Experiments-3/code-llama3p2_3b-replay-lam0p1

Reinforcement LearningLicense: apache-2.0
Tracked by OpenModelStats since October 1, 2026View on Hugging Face ↗
Downloads 30D
0
Total downloads
0
Gained 7D
—
Gained 30D
—
Likes
0
Rank · tracked models
—
–this week
Parameters
—
Last source update
Oct 5, 2026

Downloads over time

Daily gain between consecutive observations

Likes over time

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-10-01. Charts appear as daily observations accumulate.

Overview

code-llama3p2_3b-replay-lam0p1 is a reinforcement learning model published by RL-Forgetting-Experiments-3. OpenModelStats has tracked the model since Oct 1, 2026, most recently observing it Oct 6, 2026.

Published
Oct 1, 2026
Last updated
Oct 5, 2026
Library
—
License
apache-2.0
Task
Reinforcement Learning
Parameters
—
Spaces
—
Derivative models
—
Velocity 7D/day
—
safetensorsreinforcement-learninggrpocode-generationmbppcatastrophic-forgettingbase_model:meta-llama/Llama-3.2-3B-Instructbase_model:finetune:meta-llama/Llama-3.2-3B-Instructlicense:apache-2.0region:us

Publisher

RL-Forgetting-Experiments-3View publisher statistics →

Derived from meta-llama/Llama-3.2-3B-Instruct (reported by the source; parent not yet tracked).

Similar models

Models similar to code-llama3p2_3b-replay-lam0p1
ModelDL 30D
ppo-seals-CartPole-v0HumanCompatibleAI171.8K
ppo-Pendulum-v1HumanCompatibleAI46K
Tifa-Deepsex-14b-CoT-GGUFmradermacher20.9K
NoveltyAdapt-policieshelenlu9420.7K
sac-BipedalWalkerHardcore-v3sb38,773
Qwen3-14B-ARPO-DeepSearch-i1-GGUFmradermacher7,238