MATH-Llama-3.2-3B-SimKO-TS2-nosquare
xzybit/MATH-Llama-3.2-3B-SimKO-TS2-nosquare
Tracked by OpenModelStats since September 21, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 7
- Total downloadsCumulative all-time download counter reported by the source platform.
- 35
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- —
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 1
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads.
- —
- –this week
- Parameters
- 3.6B
- Last source update
- May 15, 2026
Downloads over time
Rolling recent-window downloads as reportedLikes over time
Rank among tracked models
Lower is betterNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-09-21. Charts appear as daily observations accumulate.
Overview
MATH-Llama-3.2-3B-SimKO-TS2-nosquare is a reinforcement learning model published by xzybit. OpenModelStats has tracked the model since Sep 21, 2026, most recently observing it Sep 28, 2026.
- Published
- May 15, 2026
- Last updated
- May 15, 2026
- Library
- —
- License
- llama3.2
- Task
- Reinforcement Learning
- Parameters
- 3,606,752,256
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- —
safetensorsllamamathreasoningsimkoreinforcement-learningpass@kbase_model:meta-llama/Llama-3.2-3B-Instructbase_model:finetune:meta-llama/Llama-3.2-3B-Instructlicense:llama3.2region:us
Publisher
xzybitView publisher statistics →Derived from meta-llama/Llama-3.2-3B-Instruct (reported by the source; parent not yet tracked).
Similar models
| Model | DL 30D |
|---|---|
| Qwen3-4B-Instruct-2507.prestar-RL.reason-only.lr7e-7-kl0.step-2176JackHsieh | 334 |
| VFIG-4BXunmeiLiu | 300 |
| OphVLM-R1QiZishi | 3,003 |
| beaver-7b-v1.0-costPKU-Alignment | 2,842 |
| beaver-7b-v1.0-rewardPKU-Alignment | 2,124 |
| inf-retriever-v1-proinfly | 2,350 |