Step-Audio-R1

stepfun-ai/Step-Audio-R1

Audio Text To Text33.5B parameterstransformersLicense: apache-2.0Gated
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
Downloads 30D
58
Total downloads
2,578
Gained 7D
+7(increase)
0.3%(up) vs baseline
Gained 30D
—
Likes
144
+0 in 7D
Rank · tracked models
#54,539
379(down 379 places)this week
Parameters
33.5B
Last source update
Dec 2, 2025
Momentum
21.6
  • Step-Audio-R1 gained 7 downloads during the last seven tracked days.

Downloads over time

Cumulative total downloads observed by OpenModelStats

Likes over time

Rank among tracked models

Lower is better

Overview

Step-Audio-R1 is a audio text to text model published by stepfun-ai. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it Oct 5, 2026.

Published
Nov 18, 2025
Last updated
Dec 2, 2025
Library
transformers
License
apache-2.0
Task
Audio Text To Text
Parameters
33,487,033,600
Spaces
—
Derivative models
—
Velocity 7D/day
1
transformerssafetensorsstep_audio_2text-generationaudio-reasoningchain-of-thoughtmulti-modalstep-audio-r1audio-text-to-textcustom_codearxiv:2511.15848license:apache-2.0

Similar models

Models similar to Step-Audio-R1
ModelDL 30D
Step-Audio-R1.1stepfun-ai837
Voxtral-Small-24B-2507mistralai24.6K
acestep-transcriberACE-Step3,386
MOSS-Music-8B-InstructOpenMOSS-Team3,210
MOSS-Audio-8B-InstructOpenMOSS-Team2,942
MOSS-Music-8B-ThinkingOpenMOSS-Team1,200
stepfun-ai/Step-Audio-R1 — Downloads, Growth & Stats | OpenModelStats