Hugging Face
Models
Datasets
Spaces
Posts
Docs
Enterprise
Pricing
Log In
Sign Up
CohenQu
/
DeepSeek-R1-Distill-Qwen-7B-GRPO
like
4
Text Generation
Transformers
Safetensors
hf-cmu-collab/DeepScaleR-1.5B-Preview_on-policy_GRPO
qwen2
Generated from Trainer
trl
grpo
conversational
text-generation-inference
Inference Endpoints
arxiv:
2402.03300
Model card
Files
Files and versions
Community
Train
Deploy
Use this model
main
DeepSeek-R1-Distill-Qwen-7B-GRPO
Commit History
Training in progress, step 9
0a9e8b2
verified
CohenQu
commited on
7 days ago
Training in progress, step 9
ad52c14
verified
CohenQu
commited on
7 days ago
Training in progress, step 8
f36db2b
verified
CohenQu
commited on
7 days ago
Training in progress, step 8
e674a4d
verified
CohenQu
commited on
7 days ago
Training in progress, step 7
7179850
verified
CohenQu
commited on
7 days ago
Training in progress, step 7
77997ba
verified
CohenQu
commited on
7 days ago
Training in progress, step 6
bc53665
verified
CohenQu
commited on
7 days ago
Training in progress, step 6
28765d4
verified
CohenQu
commited on
7 days ago
Training in progress, step 5
46d52b8
verified
CohenQu
commited on
7 days ago
Training in progress, step 5
313a40c
verified
CohenQu
commited on
7 days ago
Training in progress, step 4
2ce5664
verified
CohenQu
commited on
7 days ago
Training in progress, step 4
bf475f6
verified
CohenQu
commited on
7 days ago
Training in progress, step 3
398c439
verified
CohenQu
commited on
7 days ago
Training in progress, step 3
4fa5a19
verified
CohenQu
commited on
7 days ago
Training in progress, step 2
82ad5b4
verified
CohenQu
commited on
7 days ago
Training in progress, step 2
55eb55d
verified
CohenQu
commited on
7 days ago
Training in progress, step 1
6410378
verified
CohenQu
commited on
7 days ago
Training in progress, step 1
b368e6f
verified
CohenQu
commited on
7 days ago
End of training
2d8adf3
verified
CohenQu
commited on
8 days ago
Model save
6ddf61e
verified
CohenQu
commited on
8 days ago
Training in progress, step 25
ffd8ff6
verified
CohenQu
commited on
8 days ago
Training in progress, step 24
a3f80ad
verified
CohenQu
commited on
8 days ago
Training in progress, step 22
d69b21e
verified
CohenQu
commited on
8 days ago
Training in progress, step 20
8f2cb44
verified
CohenQu
commited on
8 days ago
Training in progress, step 18
344e0c8
verified
CohenQu
commited on
8 days ago
Training in progress, step 16
7e2686c
verified
CohenQu
commited on
8 days ago
Training in progress, step 14
d59e99a
verified
CohenQu
commited on
9 days ago
Training in progress, step 12
1556ad5
verified
CohenQu
commited on
9 days ago
End of training
4b3ab82
verified
CohenQu
commited on
9 days ago
Model save
2f3bf35
verified
CohenQu
commited on
9 days ago
Training in progress, step 10
8349a49
verified
CohenQu
commited on
9 days ago
Training in progress, step 10
dfef9e8
verified
CohenQu
commited on
9 days ago
Training in progress, step 9
e0cf30b
verified
CohenQu
commited on
9 days ago
Training in progress, step 8
4c7ac98
verified
CohenQu
commited on
9 days ago
Training in progress, step 8
aee2252
verified
CohenQu
commited on
9 days ago
Training in progress, step 7
425101f
verified
CohenQu
commited on
9 days ago
Training in progress, step 6
be41817
verified
CohenQu
commited on
9 days ago
Training in progress, step 6
93a1c82
verified
CohenQu
commited on
9 days ago
Training in progress, step 5
9c16aec
verified
CohenQu
commited on
9 days ago
Training in progress, step 4
8adb8de
verified
CohenQu
commited on
9 days ago
Training in progress, step 4
c5c2905
verified
CohenQu
commited on
9 days ago
Training in progress, step 3
ae88a6f
verified
CohenQu
commited on
9 days ago
Training in progress, step 2
77b71d6
verified
CohenQu
commited on
9 days ago
Training in progress, step 2
28a68fb
verified
CohenQu
commited on
9 days ago
Training in progress, step 1
c961acf
verified
CohenQu
commited on
9 days ago
End of training
9ce0500
verified
CohenQu
commited on
9 days ago
Model save
033d750
verified
CohenQu
commited on
9 days ago
Training in progress, step 20
b24eb09
verified
CohenQu
commited on
9 days ago
Training in progress, step 19
1cf7339
verified
CohenQu
commited on
9 days ago
Training in progress, step 18
85ba2cd
verified
CohenQu
commited on
9 days ago
Previous
1
...
4
5
6
7
8
Next