Claude Sonnet 4.6

Training

Pretraining data

Training data was scraped from the public internet up to January 2026, compared to Sonnet 4.5’s July 2025 and Opus 4.6 and Opus 4.5’s August 2025.

Anthropic has not disclosed whether Sonnet 4.6 shares a base model with any earlier Claude models.

Character training and constitution

Full article: Claude’s character training.

Anthropic’s character quality metrics remained unchanged from those in the Opus 4.6 System Card.

Training against chain-of-thought

Chain-of-thought supervision affected Sonnet 4.6’s post-training due to a technical error. 1

We do not train against any chains-of-thought or activations-based monitoring, with two exceptions: some SFT data that was based on transcripts from previous models was subject to filters with chain-of-thought access, and a number of environments used for Mythos Preview had a technical error that allowed reward code to see chains-of-thought.

[…] This technical error also affected the training of Claude Opus 4.6 and Claude Sonnet 4.6.

Footnotes

  1. Alignment Risk Update: Claude Mythos Preview