arxiv:2507.08396
Beier Zhu
BeierZ
AI & ML interests
None yet
Recent Activity
liked a model 26 days ago
AaronHan/OPPO liked a model 10 months ago
AaronHan/MoSEAR upvoted a paper about 1 year ago
On the Generalization of SFT: A Reinforcement Learning Perspective with
Reward RectificationOrganizations
None yet