feedback-grpo-reasoning-sft
formalmathatepfl/feedback-grpo-reasoning-sft
Tracked by OpenModelStats since September 23, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 483
- Total downloadsCumulative all-time download counter reported by the source platform.
- 483
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- —
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 0
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads. #12086 among tracked Text Generation models.
- #42,150
- 187(down 187 places)this week
- Parameters
- 8.2B
- Last source update
- Sep 23, 2026
- It moved from #11968 to #12086 among tracked Text Generation models this week.
Downloads over time
Cumulative total downloads observed by OpenModelStatsLikes over time
Rank among tracked models
Lower is betterOverview
feedback-grpo-reasoning-sft is a text generation model published by formalmathatepfl. OpenModelStats has tracked the model since Sep 23, 2026, most recently observing it Sep 24, 2026.
- Published
- Sep 23, 2026
- Last updated
- Sep 23, 2026
- Library
- transformers
- License
- other
- Task
- Text Generation
- Parameters
- 8,190,735,360
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- —
transformerssafetensorsqwen3text-generationllama-factoryfullgenerated_from_trainerconversationalbase_model:formalmathatepfl/qwen3-sft-feedback-with-proof-repair-rl-125-unmaskedbase_model:finetune:formalmathatepfl/qwen3-sft-feedback-with-proof-repair-rl-125-unmaskedlicense:othertext-generation-inference
Publisher
formalmathatepflView publisher statistics →Derived from formalmathatepfl/qwen3-sft-feedback-with-proof-repair-rl-125-unmasked (reported by the source; parent not yet tracked).
Similar models
| Model | DL 30D |
|---|---|
| Qwen3-8BQwen | 10.1M |
| Qwen3-8B-AWQQwen | 2.3M |
| Qwen3-8B-BaseQwen | 749.4K |
| DeepSeek-R1-0528-Qwen3-8Bdeepseek-ai | 682.2K |
| DeepSeek-R1-0528-Qwen3-8B-MLX-4bitlmstudio-community | 264.6K |
| DeepSeek-R1-0528-Qwen3-8B-MLX-8bitlmstudio-community | 243.9K |