<p>Anthropic’s CEO says that safety hinges on understanding how AI “thinks.” So far the evidence is disturbing.</p>