Opus 4.6 used significantly more tokens than Opus 4.5: it used 58M output tokens in adaptive thinking mode to run our Intelligence Index evaluations, ~2x Opus 4.5 (Thinking, 29M) but significantly less than GPT-5.2 with xhigh reasoning effort (130M).
No discussion yet. Be the first to share your thoughts!