Seeing this new internal model solve open after open math problem shortly after training commenced was the wildest thing I have ever witnessed at my time at OpenAI
Possibly. But tweets like this one from someone on the RL team appear to indicate that this is something weird. Would certainly be nice if OpenAI gave some more details!
This resignation seems to have gone way more viral than I would have expected. (almost 300k likes??)
Anyone have any idea why this is