Fable seems much more opinionated than previous models, but I haven’t noticed it being more deceptive. I haven’t really used it for research since Opus is already good enough at that for my purposes though[1]. When coding, Fable is much more likely to mention if tests don’t make any sense.
The only problem Opus seems to have with research is sometimes focusing on the wrong thing, but that’s usually obvious and I can just re-run it with more detail on what I’m looking for or not looking for.
Fable seems much more opinionated than previous models, but I haven’t noticed it being more deceptive. I haven’t really used it for research since Opus is already good enough at that for my purposes though[1]. When coding, Fable is much more likely to mention if tests don’t make any sense.
The only problem Opus seems to have with research is sometimes focusing on the wrong thing, but that’s usually obvious and I can just re-run it with more detail on what I’m looking for or not looking for.