Register and share your invite link to earn from video plays and referrals.

Graham Neubig
@gneubig
Associate professor @LTIatCMU. Co-founder/chief scientist @OpenHandsDev. I mostly work on modeling language.
Joined September 2010
801 Following    47.1K Followers
RL helps models learn how to reason with different strategies, but some strategies are more effective than others. But are the strategies learned by RL the ones that are most effective in improving accuracy? Our new work finds that the answer is "not always"!
Show more