But even that claim comes from the same chain-of-thought output that Anthropic says is faithful only about 25% of the time. The company is simultaneously asking the public to trust Claude’s stated reasoning. while publishing research saying you shouldn’t.

Reasoning models don't always say what they thinkResearch from Anthropic on the faithfulness of AI models' Chain-of-Thoughtwww.anthropic.com