a3-rl-DCAgent_exp_rpt_e2egit-v2-10-8B

laion/a3-rl-DCAgent_exp_rpt_e2egit-v2-10-8B

Reinforcement Learning8.2B parametersLicense: apache-2.0
Tracked by OpenModelStats since August 29, 2026View on Hugging Face ↗
Downloads 30D
1,082
Total downloads
1,190
Gained 7D
Gained 30D
Likes
0
Rank · tracked models
this week
Parameters
8.2B
Last source update
May 26, 2026

Downloads over time

Cumulative total downloads observed by OpenModelStats

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-29. Charts appear as daily observations accumulate.

Likes over time

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-29. Charts appear as daily observations accumulate.

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-29. Charts appear as daily observations accumulate.

Overview

a3-rl-DCAgent_exp_rpt_e2egit-v2-10-8B is a reinforcement learning model published by laion. OpenModelStats has tracked the model since Aug 29, 2026, most recently observing it 3h ago.

Published
May 26, 2026
Last updated
May 26, 2026
Library
License
apache-2.0
Task
Reinforcement Learning
Parameters
8,190,735,360
Spaces
Derivative models
Velocity 7D/day
safetensorsqwen3reinforcement-learningrlskyrlterminal-benchagentbase_model:laion/GLM-4_7-swesmith-sandboxes-with_tests-oracle_verified_120s-maxeps-131k-fixthinkbase_model:finetune:laion/GLM-4_7-swesmith-sandboxes-with_tests-oracle_verified_120s-maxeps-131k-fixthinklicense:apache-2.0region:us

Publisher

laionView publisher statistics →

Derived from laion/GLM-4_7-swesmith-sandboxes-with_tests-oracle_verified_120s-maxeps-131k-fixthink (reported by the source; parent not yet tracked).

Similar models

Models similar to a3-rl-DCAgent_exp_rpt_e2egit-v2-10-8B
ModelDL 30D
MetaAgent-XMercury73532,866
VisualQuality-R1-7BTianheWu10.4K
DiBO-TFBind10zpointsun701
inf-retriever-v1-proinfly2,003
beaver-7b-v1.0-costPKU-Alignment6,284
beaver-7b-v1.0-rewardPKU-Alignment5,333