General-purpose competency isn’t just a function of the model alone. It depends on harnesses, and harnesses can be developed and improved over time. No one really knows what level current public models could reach with the right harness development, let alone a new model.
I don’t expect this thing to probably take over either. But if there is an AI takeover I expect to be surprised by it, and expect normalization of deviance to let people chug along right up to the moment.
My actual main reason for hope that an AI takeover ultimately doesn’t happen is, I think, probably actually yours too given your previous discussion of intent-aligned AI? That it’s programmed to follow user instructions, and maybe OpenAI will come to their senses and train it make sure that apparent instructions are grounded in original user intent and not just harness-internal AI chatter. (as well as overall safety guidelines, I hope, though the commercial incentive may be to weaken those)
I think OpenAI’s stated belief on the reason is most likely: