Anthropic’s Jackson Kernion has explained why Claude’s prose became noticeably worse after Opus 4.6. The cause: training shifted toward math, code, and technical explanations aimed at other language models. LLMs have larger working memory than humans and pick up on fine-grained details more precisely, so the writing style optimized for them reads to people as dense, impenetrable slabs of text. The culprit sits in the RL reward signals: some rewards incentivize text that is clear to the model, others text that is clear to humans. The more math and code in training, the harder it becomes to reward the simpler explanations a person can actually read. Kernion says Opus 5.5 has finally found a better balance, though he notes the problem is difficult and work continues.

X: Jackson Kernion

Related: Anthropic Reduces Claude’s Flattery in Relationship Advice, Anthropic Releases Claude Opus 5.5