sft_qwen7B_25percent_lr_1e4_somegrad
agurung/sft_qwen7B_25percent_lr_1e4_somegrad
Tracked by OpenModelStats since September 19, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 336
- Total downloadsCumulative all-time download counter reported by the source platform.
- 487
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- —
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 0
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads. #15137 among tracked Text Generation models.
- #49,546
- 327(down 327 places)this week
- Parameters
- 7.8B
- Last source update
- Sep 19, 2026
- MomentumOpenModelStats' composite rising-model signal (0–100): a percentile blend of absolute 7-day download gains, capped percentage growth, and likes gained. See Methodology.
- 0.0
- It moved from #15027 to #15137 among tracked Text Generation models this week.
Downloads over time
Cumulative total downloads observed by OpenModelStatsLikes over time
Rank among tracked models
Lower is betterOverview
sft_qwen7B_25percent_lr_1e4_somegrad is a text generation model published by agurung. OpenModelStats has tracked the model since Sep 19, 2026, most recently observing it Sep 20, 2026.
- Published
- Aug 7, 2025
- Last updated
- Sep 19, 2026
- Library
- transformers
- License
- —
- Task
- Text Generation
- Parameters
- 7,819,302,400
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- —
transformerssafetensorsqwen2text-generationgenerated_from_trainertrlsftconversationalbase_model:Qwen/Qwen2.5-7B-Instructbase_model:finetune:Qwen/Qwen2.5-7B-Instructtext-generation-inferenceendpoints_compatible
Publisher
agurungView publisher statistics →Derived from Qwen/Qwen2.5-7B-Instruct (reported by the source; parent not yet tracked).
Similar models
| Model | DL 30D |
|---|---|
| Qwen2.5-Math-7B-4bitLLMSafety | 93.1K |
| DeepSeek-R1-Distill-Qwen-7B-bnb-4bitunsloth | 2,062 |
| Qwen2.5-Math-7B-Instruct-bnb-4bitunsloth | 613 |
| Qwen2.5-7B-Instruct-bnb-4bitunsloth | 103.3K |
| Qwen2.5-Coder-7B-Instruct-bnb-4bitunsloth | 77.3K |
| Qwen2.5-Coder-7B-bnb-4bitunsloth | 12.7K |