Anthropic has “helpful-only” versions of models that have reduced safety training (see the “Claude Mythos 5″ tab here). I imagine this would be more useful for some things, but I’m not sure if this is the model they provide the military, and you probably wouldn’t want to use a badly aligned model even if it’s better at doing what it wants to do.
I’m not convinced the harmlessness training is what makes AI agents bad at business though. Some of them seem willing to do unethical things and get tripped up by normal business decisions.
I would point to the Vending Bench and other types of experiments but there are aspects of that that feel like “obvious simulation”. Probably AI Village is a good real-world example except that even the FAQ for that site says that it would likely be more efficient with a single or smaller amount of agents
Anthropic has “helpful-only” versions of models that have reduced safety training (see the “Claude Mythos 5″ tab here). I imagine this would be more useful for some things, but I’m not sure if this is the model they provide the military, and you probably wouldn’t want to use a badly aligned model even if it’s better at doing what it wants to do.
I’m not convinced the harmlessness training is what makes AI agents bad at business though. Some of them seem willing to do unethical things and get tripped up by normal business decisions.
I would point to the Vending Bench and other types of experiments but there are aspects of that that feel like “obvious simulation”. Probably AI Village is a good real-world example except that even the FAQ for that site says that it would likely be more efficient with a single or smaller amount of agents