GRPO-Think-7B-8k

Aletheia-Bench/GRPO-Think-7B-8k

Text Generation7.6B parameterstransformersLicense: cc-by-nc-sa-4.0
Tracked by OpenModelStats since August 25, 2026View on Hugging Face ↗
Downloads 30D
274
Total downloads
359
Gained 7D
Gained 30D
Likes
0
Rank · tracked models
this week
Parameters
7.6B
Last source update
Aug 25, 2026

Downloads over time

Rolling recent-window downloads as reported

Likes over time

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-25. Charts appear as daily observations accumulate.

Overview

GRPO-Think-7B-8k is a text generation model published by Aletheia-Bench. OpenModelStats has tracked the model since Aug 25, 2026, most recently observing it 26h ago.

Published
Dec 12, 2025
Last updated
Aug 25, 2026
Library
transformers
License
cc-by-nc-sa-4.0
Task
Text Generation
Parameters
7,615,616,512
Spaces
0
Derivative models
Velocity 7D/day
transformerssafetensorsqwen2text-generationcodecode-verificationrlvrgrpoconversationaldataset:Aletheia-Bench/Aletheia-Trainarxiv:2601.12186base_model:deepseek-ai/DeepSeek-R1-Distill-Qwen-7B

Model family

Based on base-model relationships reported by the source's metadata.

Similar models

Models similar to GRPO-Think-7B-8k
ModelDL 30D
Qwen2.5-7B-InstructQwen10.8M
Qwen2.5-7B-Instruct-AWQQwen4M
Qwen2.5-Coder-7B-InstructQwen2.4M
Qwen2.5-Coder-7B-Instruct-AWQOrion-zhen590.2K
Qwen2.5-7BQwen577.3K
Qwen2.5-Coder-7BQwen575.6K