Newer models can cost more per word at the same price per token
Claude Opus 4.7 and later use a newer tokenizer that produces roughly 30% more tokens for the same text than Claude Sonnet 4.6 and earlier. A model with an identical headline price can therefore cost about 30% more for identical input. Cost models that compare only $/MTok miss this entirely.