teaching-to-reward-hack
wiruka/teaching-to-reward-hack
peft
Tracked by OpenModelStats since October 3, 2026View on Hugging Face ↗
- Downloads 30DRolling recent-window download count as reported by the source platform.
- 0
- Total downloadsCumulative all-time download counter reported by the source platform.
- 0
- Gained 7DIncrease in cumulative total downloads between OpenModelStats' own observations seven days apart.
- —
- Gained 30DIncrease in cumulative total downloads between OpenModelStats' own observations roughly thirty days apart.
- —
- LikesCommunity likes reported by the source platform.
- 0
- Rank · tracked modelsPosition among all eligible models tracked by OpenModelStats, ordered by recent downloads.
- —
- –this week
- Parameters
- —
- Last source update
- Oct 2, 2026
Downloads over time
Cumulative total downloads observed by OpenModelStatsLikes over time
Rank among tracked models
Lower is betterNot enough tracked history yet for a chart.
OpenModelStats began tracking this model on 2026-10-03. Charts appear as daily observations accumulate.
Overview
teaching-to-reward-hack is an AI model published by wiruka. OpenModelStats has tracked the model since Oct 3, 2026, most recently observing it Oct 4, 2026.
- Published
- Oct 2, 2026
- Last updated
- Oct 2, 2026
- Library
- peft
- License
- —
- Task
- —
- Parameters
- —
- Spaces
- —
- Derivative models
- —
- Velocity 7D/day
- —
peftsafetensorsloratinkerbase_model:openai/gpt-oss-120bbase_model:adapter:openai/gpt-oss-120bregion:us
Publisher
wirukaView publisher statistics →Derived from openai/gpt-oss-120b (reported by the source; parent not yet tracked).
Similar models
| Model | DL 30D |
|---|---|
| all-MiniLM-L6-v2sentence-transformers | 235.5M |
| ms-marco-MiniLM-L6-v2cross-encoder | 83.8M |
| bge-small-en-v1.5BAAI | 62.6M |
| paraphrase-multilingual-MiniLM-L12-v2sentence-transformers | 50.7M |
| electra-base-discriminatorgoogle | 45.9M |
| bert-base-uncasedgoogle-bert | 38.9M |