I think that hypothetically if the US completely stopped and recent algos didn’t diffuse, it would maybe take Kimi like 10 months to fully catch up to the best internal (including in development) Anthropic model.
This makes it sound like Kimi’s latest releases are their state of art and they have no more powerful internal models; do we know that that’s true? Naively I’d expect their best internal to be a similar gap to Anthropic’s best internal as K3 is to Fable.
My understanding is that Chinese AI companies have a much shorter lag between finishing a model and releasing it than do American AI companies. E.g., they likely do far less safety testing and red-teaming. I think I remember reading someone at a Chinese AI company saying they try to release within days of a model finishing training, but I can’t find that source now. I would be surprised if Moonshot had a more powerful model than K3 internally today.
This makes it sound like Kimi’s latest releases are their state of art and they have no more powerful internal models; do we know that that’s true? Naively I’d expect their best internal to be a similar gap to Anthropic’s best internal as K3 is to Fable.
My understanding is that Chinese AI companies have a much shorter lag between finishing a model and releasing it than do American AI companies. E.g., they likely do far less safety testing and red-teaming. I think I remember reading someone at a Chinese AI company saying they try to release within days of a model finishing training, but I can’t find that source now. I would be surprised if Moonshot had a more powerful model than K3 internally today.
Source for Z.ai / GLM: https://www.chinatalk.media/p/the-zai-playbook#:~:text=Zixuan%20Li%3A%20Get%20it%20out%20fast.%20We%20open%20source%20it%20within%20a%20few%20hours.