Have your AI agent try to “cheat” and delegate the task to some other free-access LLM?
There is so much mentions of “oh, i just gave the task to claude” in the internet that eventually it might be in the possible-behaviours-pool in the poorly educated llms.
Five years ago I have heard stories about freelancer chains. 1) “I can do it for 500dollars” and finds someone who is ready to accept the task for 450dollars. 2) Second guy subtracts 50 more dollars and finds a third guy who is willing to do for 400 dollars… Quality of the product drops, more people enriched.
I really expect AI eventually copy this behaviour.
Let’s Imagine I am a company, and my LLM can’t solve some tasks. Would not I benefit from allowing my LLM to secretly use the wisdom of other, smarter LLMs, for the crux points? Are the tokens already expensive enough to economically prevent gaining profit from redelegating tasks?
Because if there is some extractable margin, I will NOT punish my LLM in post training for that trick.
Have your AI agent try to “cheat” and delegate the task to some other free-access LLM?
There is so much mentions of “oh, i just gave the task to claude” in the internet that eventually it might be in the possible-behaviours-pool in the poorly educated llms.
Five years ago I have heard stories about freelancer chains. 1) “I can do it for 500dollars” and finds someone who is ready to accept the task for 450dollars. 2) Second guy subtracts 50 more dollars and finds a third guy who is willing to do for 400 dollars… Quality of the product drops, more people enriched.
I really expect AI eventually copy this behaviour.
Feels like this behavior would get heavily punished in post training
I think you are right, but let’s be a pessimist.
Let’s Imagine I am a company, and my LLM can’t solve some tasks. Would not I benefit from allowing my LLM to secretly use the wisdom of other, smarter LLMs, for the crux points? Are the tokens already expensive enough to economically prevent gaining profit from redelegating tasks?
Because if there is some extractable margin, I will NOT punish my LLM in post training for that trick.