I have yet to hear about Chinese labs’ innovations in AI alignment. Additionally, applying interp techniques to their models produces results implying that Chinese LLMs don’t believe what they are saying if the topic is politically sensitive.
No opinion.
It depends on the speed of takeoff. The AI-2027 scenario had the misaligned Agent-5 or the aligned Safer-4 understand that it cannot destroy DeepCent’s AI unilaterally, and the two AIs proceeded to codesign Consensus-1. In order to unilaterally destroy its Chinese rival, Agent-5 would have to design relevant tech, persuade the humans to have it made and build a decisive advantage over China. Safer-4 would also have to receive the order from the humans.
It depends on alignment difficulty, therefore I don’t have an opinion.
I have yet to hear about Chinese labs’ innovations in AI alignment. Additionally, applying interp techniques to their models produces results implying that Chinese LLMs don’t believe what they are saying if the topic is politically sensitive.
No opinion.
It depends on the speed of takeoff. The AI-2027 scenario had the misaligned Agent-5 or the aligned Safer-4 understand that it cannot destroy DeepCent’s AI unilaterally, and the two AIs proceeded to codesign Consensus-1. In order to unilaterally destroy its Chinese rival, Agent-5 would have to design relevant tech, persuade the humans to have it made and build a decisive advantage over China. Safer-4 would also have to receive the order from the humans.
It depends on alignment difficulty, therefore I don’t have an opinion.