I mainly want to see what strategy they used out of curiosity. I don’t think they are just running RLHF on whatever deep-learning based super efficient learner novel architecture they cooked up. Would be interesting to see!
If there isn’t a superefficient novel architecture, then Sutskever is to be arresred for wholesale fraud. If there is, then Sutskever’s startup makes the world LESS safe to live in, and Sutskever failed at bringing about his promises. How could one arrange for the potential architecture to be leaked into Anthropic (or, at least, OAI/GDM) for thorough analysis?
I mainly want to see what strategy they used out of curiosity. I don’t think they are just running RLHF on whatever deep-learning based super efficient learner novel architecture they cooked up. Would be interesting to see!
If there isn’t a superefficient novel architecture, then Sutskever is to be arresred for wholesale fraud. If there is, then Sutskever’s startup makes the world LESS safe to live in, and Sutskever failed at bringing about his promises. How could one arrange for the potential architecture to be leaked into Anthropic (or, at least, OAI/GDM) for thorough analysis?