This collection contains curriculum-RLed Olmo models.
SeanWang0027 PRO
SeanWang0027
AI & ML interests
LLM Post-Training
Recent Activity
authored a paper about 19 hours ago
Learning from Teacher Continuations at Student States upvoted a paper 1 day ago
Learning from Teacher Continuations at Student States upvoted a paper 1 day ago
Selecting Diverse SFT Traces Improves Post-RL Generalization