I went ahead tried a run with “xhigh”, which I think OpenRouter will translate to Anthropic’s “max”. Didn’t seem to make a lot of difference. It didn’t fixate as much, but it mostly seems to have just gotten more conservative about what it was willing to say, and it didn’t catch anything new.
Generally I’ve found that when models miss hints at the beginning, I have to steer them pretty hard conversationally if I want them to see those hints, so I’m guessing that more self-talk probably won’t be helpful for most of them.
I went ahead tried a run with “xhigh”, which I think OpenRouter will translate to Anthropic’s “max”. Didn’t seem to make a lot of difference. It didn’t fixate as much, but it mostly seems to have just gotten more conservative about what it was willing to say, and it didn’t catch anything new.
Generally I’ve found that when models miss hints at the beginning, I have to steer them pretty hard conversationally if I want them to see those hints, so I’m guessing that more self-talk probably won’t be helpful for most of them.