models also out persuade national championships debaters and professional canvassers, although the models regress when their ‘quantity of facts presented’ is limited to the human norms.
The study the authors link indicates that Opus 4.1 from a year ago is marginally better than 4o and those both are way better than Opus 4.6 and GPT-5.4 from this year. If we believe these results, this doesn’t demonstrate any progress towards the “superpersuasion” but quite to the contrary
Doesn’t say how the human persuaders were selected, and if they weren’t top experts it doesn’t demonstrate superhuman ability.
models also out persuade national championships debaters and professional canvassers, although the models regress when their ‘quantity of facts presented’ is limited to the human norms.
It’s covered in the first dispatch here.
https://aistop.watch/p/shifting-perspectives
The study the authors link indicates that Opus 4.1 from a year ago is marginally better than 4o and those both are way better than Opus 4.6 and GPT-5.4 from this year. If we believe these results, this doesn’t demonstrate any progress towards the “superpersuasion” but quite to the contrary