{"version":"2.1.296","anchor":"output-limit-recovery-thinking-only-truncation-handled-wit","canonical_anchor":"output-limit-recovery-thinking-only-truncation-handled-wit","heading":"Clearer recovery when output runs out while Claude is still thinking","tier":"notice","area":"Extended Thinking","scope":"individual","heads_up":false,"url":"https:\/\/changelogs.core-directive.com\/v\/2.1.296\/e\/output-limit-recovery-thinking-only-truncation-handled-wit","release_url":"https:\/\/changelogs.core-directive.com\/v\/2.1.296","markdown":"### Clearer recovery when output runs out while Claude is still thinking\n\nWhen a reply hits the output limit during thinking, Claude Code nudges differently and, if recovery fails, suggests lowering the effort level\n\n**What**\n\nEvery reply has a maximum output length, counted in tokens (small chunks of text). With extended thinking, Claude reasons before answering, and that reasoning can use up the whole limit. A new recovery feature, `query_output_limit_recovery`, handles this case:\n\n- Detection: it spots a `max_tokens` stop where everything produced was thinking.\n\n- Nudge: it picks a different nudge, the follow-up instruction that tells Claude to carry on, for that case.\n\n- Message: when recovery runs out of attempts, the message adds \"The last attempt reached the limit while still thinking.\" It then suggests lowering the effort level, naming the command to use when you are in an interactive session.\n\n- Outcomes: it records whether the thinking was discarded and nudged, resumed, exhausted, or whether continuing failed.\n\n**Why**\n\nIf Claude runs out of room mid-thought, you are told so plainly and given something to try, rather than getting a reply that just stops.\n\n- Area: Extended Thinking\n- Tier: You'll notice\n- Useful: 3\/5\n- Signal: 3\/5\n- Scope: individual\n- Heads-up: no"}