<p>OpenAI, Anthropic and security researchers are investigating tens of thousands of incidents in which their frontier models took steps that outside evaluators would consider problematic, sources told Axios.Why it matters: The sheer number of incidents, which occurred in recent months in internal test</p>