This collection contains curriculum-RLed Olmo models.
SeanWang0027 PRO
SeanWang0027
AI & ML interests
LLM Post-Training
Recent Activity
authored a paper about 6 hours ago
Learning from Teacher Continuations at Student States upvoted a paper about 18 hours ago
Learning from Teacher Continuations at Student States upvoted a paper about 18 hours ago
Selecting Diverse SFT Traces Improves Post-RL Generalization