Solving a metagame from first principles, without guides or hundreds of hours sounds like a superhuman skill. Which I agree current LLMs lack, but I bring up fairness because if we’re listing “not superhuman” as a limitation of current LLMs then we might as well complain they can’t prove the Riemann hypothesis.
I have been very impressed with Fable, so I’m making a note to self to spend a bunch of tokens having it learn Pokemon and see if it can put out a half-decent performance, a few months from now when the tokens are cheaper.
The US government can compel the disclosure of all recorded LLM conversations that it has a good reason to suspect contain evidence of wrongdoing, in accordance with longstanding norms about compelling companies to turn over evidence of wrongdoing. In theory, the US government can lie about its reasons for wanting access to any LLM conversation, but it is a substantial escalation that Plan A grants them hardware access to data which could previously only be obtained by lying to a third party.