Hugging Face
Models
Datasets
Spaces
Posts
Docs
Enterprise
Pricing
Log In
Sign Up
CohenQu
/
DeepSeek-R1-Distill-Qwen-7B-GRPO
like
4
Text Generation
Transformers
Safetensors
hf-cmu-collab/DeepScaleR-1.5B-Preview_on-policy_GRPO
qwen2
Generated from Trainer
trl
grpo
conversational
text-generation-inference
Inference Endpoints
arxiv:
2402.03300
Model card
Files
Files and versions
Community
Train
Deploy
Use this model
b368e6f
DeepSeek-R1-Distill-Qwen-7B-GRPO
Commit History
Training in progress, step 1
b368e6f
verified
CohenQu
commited on
7 days ago
End of training
2d8adf3
verified
CohenQu
commited on
8 days ago
Model save
6ddf61e
verified
CohenQu
commited on
8 days ago
Training in progress, step 25
ffd8ff6
verified
CohenQu
commited on
8 days ago
Training in progress, step 24
a3f80ad
verified
CohenQu
commited on
8 days ago
Training in progress, step 22
d69b21e
verified
CohenQu
commited on
8 days ago
Training in progress, step 20
8f2cb44
verified
CohenQu
commited on
8 days ago
Training in progress, step 18
344e0c8
verified
CohenQu
commited on
8 days ago
Training in progress, step 16
7e2686c
verified
CohenQu
commited on
8 days ago
Training in progress, step 14
d59e99a
verified
CohenQu
commited on
8 days ago
Training in progress, step 12
1556ad5
verified
CohenQu
commited on
8 days ago
End of training
4b3ab82
verified
CohenQu
commited on
8 days ago
Model save
2f3bf35
verified
CohenQu
commited on
8 days ago
Training in progress, step 10
8349a49
verified
CohenQu
commited on
8 days ago
Training in progress, step 10
dfef9e8
verified
CohenQu
commited on
8 days ago
Training in progress, step 9
e0cf30b
verified
CohenQu
commited on
8 days ago
Training in progress, step 8
4c7ac98
verified
CohenQu
commited on
8 days ago
Training in progress, step 8
aee2252
verified
CohenQu
commited on
8 days ago
Training in progress, step 7
425101f
verified
CohenQu
commited on
8 days ago
Training in progress, step 6
be41817
verified
CohenQu
commited on
8 days ago
Training in progress, step 6
93a1c82
verified
CohenQu
commited on
8 days ago
Training in progress, step 5
9c16aec
verified
CohenQu
commited on
8 days ago
Training in progress, step 4
8adb8de
verified
CohenQu
commited on
8 days ago
Training in progress, step 4
c5c2905
verified
CohenQu
commited on
8 days ago
Training in progress, step 3
ae88a6f
verified
CohenQu
commited on
8 days ago
Training in progress, step 2
77b71d6
verified
CohenQu
commited on
8 days ago
Training in progress, step 2
28a68fb
verified
CohenQu
commited on
8 days ago
Training in progress, step 1
c961acf
verified
CohenQu
commited on
8 days ago
End of training
9ce0500
verified
CohenQu
commited on
9 days ago
Model save
033d750
verified
CohenQu
commited on
9 days ago
Training in progress, step 20
b24eb09
verified
CohenQu
commited on
9 days ago
Training in progress, step 19
1cf7339
verified
CohenQu
commited on
9 days ago
Training in progress, step 18
85ba2cd
verified
CohenQu
commited on
9 days ago
Training in progress, step 17
88632ae
verified
CohenQu
commited on
9 days ago
Training in progress, step 16
9052815
verified
CohenQu
commited on
9 days ago
Training in progress, step 15
76337b7
verified
CohenQu
commited on
9 days ago
Training in progress, step 14
3cf3755
verified
CohenQu
commited on
9 days ago
Training in progress, step 13
9acac60
verified
CohenQu
commited on
9 days ago
Training in progress, step 12
2dd44e3
verified
CohenQu
commited on
9 days ago
Training in progress, step 11
5900317
verified
CohenQu
commited on
9 days ago
Training in progress, step 9
cbeff76
verified
CohenQu
commited on
9 days ago
Training in progress, step 8
473a5b2
verified
CohenQu
commited on
9 days ago
Training in progress, step 7
e774e36
verified
CohenQu
commited on
9 days ago
Training in progress, step 6
0ae729a
verified
CohenQu
commited on
9 days ago
Training in progress, step 5
5da45db
verified
CohenQu
commited on
9 days ago
Training in progress, step 4
91f02da
verified
CohenQu
commited on
9 days ago
Training in progress, step 3
4f28678
verified
CohenQu
commited on
9 days ago
Training in progress, step 2
fff582c
verified
CohenQu
commited on
9 days ago
Training in progress, step 1
b7cd301
verified
CohenQu
commited on
9 days ago
End of training
6140715
verified
CohenQu
commited on
10 days ago
Previous
1
2
3
Next