Qwen3.5-4B-LoRA-GRPO-CyberSec-Reasoner

sajan-sarker/Qwen3.5-4B-LoRA-GRPO-CyberSec-Reasoner

Text Generation4.2B parameterstransformers
Tracked by OpenModelStats since August 24, 2026View on Hugging Face ↗
Downloads 30D
80
Total downloads
450
Gained 7D
Gained 30D
Likes
2
Rank · tracked models
this week
Parameters
4.2B
Last source update
Apr 18, 2026

Downloads over time

Cumulative total downloads observed by OpenModelStats

Likes over time

Rank among tracked models

Lower is better

Not enough tracked history yet for a chart.

OpenModelStats began tracking this model on 2026-08-24. Charts appear as daily observations accumulate.

Overview

Qwen3.5-4B-LoRA-GRPO-CyberSec-Reasoner is a text generation model published by sajan-sarker. OpenModelStats has tracked the model since Aug 24, 2026, most recently observing it 27h ago.

Published
Apr 17, 2026
Last updated
Apr 18, 2026
Library
transformers
License
Task
Text Generation
Parameters
4,205,751,296
Spaces
0
Derivative models
Velocity 7D/day
transformerssafetensorsqwen3_5_texttext-generationcybersecuritygrporeasoningtrlvulnerability-analysiscve-to-cwe mappingconversationalarxiv:1910.09700

Model family

Based on base-model relationships reported by the source's metadata.

Similar models

Models similar to Qwen3.5-4B-LoRA-GRPO-CyberSec-Reasoner
ModelDL 30D
zira-researcher0xvoid000012.3K
Agents-A1-4B-Fable-Preview-heretichotdogs2,447
Polaris-V1nitrai-research1,981
Qwen3.5-4B-Safety-ThinkingMerlin-Research1,534
Qwen3.5-4B-MATH-ReAct-Agentic-ESOptzz1358m486
Qwen3.5-4B-abliteratedwangzhang464