Radna
radna
AI & ML interests
None yet
Organizations
When ReCAP-32B-Instruct?
1
#2 opened 2 months ago
by
radna
When ReCAP-32B-Instruct?
1
#1 opened 2 months ago
by
radna
QwQ-32B
10
#8 opened over 1 year ago
by
sm54
Quant settings?
3
#2 opened over 1 year ago
by
radna
[Experiment] Applying GRPO to DeepSeek-R1-Distill-Qwen-1.5B with LIMO
😎🔥 22
22
#15 opened over 1 year ago
by
lewtun
training code
2
#1 opened over 1 year ago
by
Ping404
The inference performance of the DeepSeek-R1-AWQ model is weak compared to the DeepSeek-R1 model
👍 3
8
#3 opened over 1 year ago
by
qingqingz916
Fix task tag
#1 opened almost 2 years ago
by
merve