What’s Trending in AI?

Verifier Errors in RLVR: Reward Hacking, Limits of Feedback, and Selective Control · Bharat Hunt