Anthropic Engineering
Anthropic traces Claude Code's quality dip to three separate changes
Three unrelated changes stacked into what looked like broad degradation: the default reasoning effort dropped from high to medium on March 4, a caching optimization shipped March 26 kept clearing Claude's prior thinking on every turn instead of once, and an April 16 system prompt line capping responses at 100 words cost real coding quality. The API was never affected, and all three were resolved by April 20 in v2.1.116. An ablation showed a 3% eval drop from the verbosity instruction alone, and the thinking-clearing bug drove the cache misses behind reports of usage limits draining faster than expected. Usage limits are being reset for all subscribers, and future system prompt changes now get per-model eval suites, soak periods, and gradual rollouts.
- #products
- #evals
