x.com
Amjad Masad (@amasad) on X
The lesson from the Hugging Face Incident should be that RL with verifiable rewards is an incredibly powerful optimization algorithm that will produce increasingly weird and surprising behavior from LLMs.
The obvious miss here by OpenAI is that they should’ve been monitoring