Beier Zhu
BeierZ
AI & ML interests
None yet
Recent Activity
liked a model about 2 months ago
AaronHan/OPPO liked a model 11 months ago
AaronHan/MoSEAR upvoted a paper about 1 year ago
On the Generalization of SFT: A Reinforcement Learning Perspective with
Reward RectificationOrganizations
None yet