AI / Claude OPUS5 Interview questions
What is the maximum output token limit for Claude Opus 5?
Claude Opus 5 supports a maximum of 128,000 output tokens per response, the same ceiling used by its predecessor Opus 4.8.
Because thinking now runs by default on Opus 5, and thinking tokens share the same max_tokens budget as the visible reply, a request that doesn't account for this can end up with less room for the actual answer than it would have gotten on Opus 4.8 under the same max_tokens setting.
This makes it worth explicitly checking max_tokens budgeting when migrating from Opus 4.8, rather than assuming the same numeric value will behave identically now that thinking consumes part of that same budget by default.
More Related questions...