rl-training-debug-artifacts

prism-drift/rl-training-debug-artifacts

Reinforcement LearningpeftLicense: apache-2.0
Tracked by OpenModelStats since August 26, 2026View on Hugging Face ↗
Downloads 30D
0
Total downloads
0
Gained 7D
Gained 30D
Likes
0
Rank · tracked models
this week
Parameters
Last source update
Aug 26, 2026

Downloads over time

Rolling recent-window downloads as reported

Likes over time

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-26. Charts appear as daily observations accumulate.

Overview

rl-training-debug-artifacts is a reinforcement learning model published by prism-drift. OpenModelStats has tracked the model since Aug 26, 2026, most recently observing it 17h ago.

Published
Aug 26, 2026
Last updated
Aug 26, 2026
Library
peft
License
apache-2.0
Task
Reinforcement Learning
Parameters
Spaces
Derivative models
Velocity 7D/day
peftsafetensorsqwen3.5reinforcement-learningdpogrpoppodebuggingbase_model:Qwen/Qwen3.5-4Bbase_model:adapter:Qwen/Qwen3.5-4Blicense:apache-2.0region:us

Publisher

prism-driftView publisher statistics →

Derived from Qwen/Qwen3.5-4B (reported by the source; parent not yet tracked).

Similar models

Models similar to rl-training-debug-artifacts
ModelDL 30D
joint-space-empowermentjamesheald229.7K
ppo-LunarLander-v3-flipKaptainKris144.5K
ppo-seals-CartPole-v0HumanCompatibleAI45.2K
ppo-Pendulum-v1HumanCompatibleAI17.1K
VisualQuality-R1-7BTianheWu8,110
beaver-7b-v1.0-costPKU-Alignment6,348