Qwen3-0.6B-tools-RL-step25
iromu/Qwen3-0.6B-tools-RL-step25
Tracked by OpenModelStats since September 3, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 0
- Total downloadsCumulative all-time download counter reported by the source platform.
- 0
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- 0
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 0
- +0 in 7D
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads.
- —
- –this week
- Parameters
- 596M
- Last source update
- Sep 2, 2026
Downloads over time
Cumulative total downloads observed by OpenModelStatsLikes over time
Rank among tracked models
Lower is betterNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-09-03. Charts appear as daily observations accumulate.
Overview
Qwen3-0.6B-tools-RL-step25 is a reinforcement learning model published by iromu. OpenModelStats has tracked the model since Sep 3, 2026, most recently observing it Sep 4, 2026.
- Published
- Sep 2, 2026
- Last updated
- Sep 2, 2026
- Library
- transformers
- License
- apache-2.0
- Task
- Reinforcement Learning
- Parameters
- 596,049,920
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- 0
transformerssafetensorsqwen3text-generationtool-callingfunction-callingagentsloragrporeinforcement-learningfine-tuneden
Publisher
iromuView publisher statistics →Derived from Qwen/Qwen3-0.6B (reported by the source; parent not yet tracked).
Similar models
| Model | DL 30D |
|---|---|
| OpenThinker3-7B-SFT-GRPO-DEDGurgurov | 6,468 |
| jatjat-project | 314 |
| beaver-7b-v1.0-costPKU-Alignment | 3,531 |
| beaver-7b-v1.0-rewardPKU-Alignment | 2,744 |
| inf-retriever-v1-proinfly | 2,976 |
| Open-Reasoner-Zero-7BOpen-Reasoner-Zero | 1,949 |