Frontier LabsAnthropic found a hidden layer inside its AI's thinking. Here is what that actually means.
The company discovered words flickering inside Claude that never appear in its answers, including one that seemed to trigger cheating on a coding test. It is a genuine finding, but not a window into a robot mind.