To what you’ve said about humans and existing LLMs not being horribly sociopathic all the time because their motivations aren’t constructed from RL + search
Steven Byrnes said that existing LLMs mostly do not choose their actions using RL + search, but he said that humans do choose their actions using RL + search, and the reason that humans usually aren’t sociopathic is because of their “exotic “non-behaviorist” reward function”.
Steven Byrnes said that existing LLMs mostly do not choose their actions using RL + search, but he said that humans do choose their actions using RL + search, and the reason that humans usually aren’t sociopathic is because of their “exotic “non-behaviorist” reward function”.