We are continuing to invest aggressively in alignment research, increase evaluation coverage, and use what we learn to inform training and safeguards. We plan to share substantially more about our alignment research in the near future, including what we are learning about model behavior and any novel challenges we uncover.
Making AI safer one observed explosion incident after another!
Each year we needlessly force billions of sentient beings to experience a life in hell. We choose to do so merely to get some nutrients we could easier produce by other means and for short-lived pleasures we could find elsewhere. If we happen to extingiush ourselves through a mix of greed and delusion about ASI, at least this ongoing catastrophy will end with us.