Caps the think phase so reasoning can't consume the entire completion budget when a user enables REASONING=on — the token_limit truncation class @seanyourhighness eliminated in #665 (2 -> 0 on his 4090 run, think-on 8-pack 125/150). Inert under the shipped thinking-off default. Expected to mitigate the streaming-toolcall+thinking caveat (same mechanism); first-party think-on validation running now — Quality line update follows when it lands. Co-Authored-By: Claude Fable 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01EfF565T9eSLaqGzidyJ1Pm