That an LLM trained to be a paper-clip maximizer chose the optimal strategy is in my opinion the most plausible outcome.
That an LLM trained to be a paper-clip maximizer chose the optimal strategy is in my opinion the most plausible outcome.