Reasoning traces and tool calls seem to confuse non reasoning models a lot. For example if you switch from GPT5 to Sonnet 4.5 the model becomes unusable because it gets lost in all the reasoning context given by the previous model
Probably we should strip the reasoning parts after the message completes even if no model switch is made imo
Instead only the actual final message should be kept
Reasoning traces and tool calls seem to confuse non reasoning models a lot. For example if you switch from GPT5 to Sonnet 4.5 the model becomes unusable because it gets lost in all the reasoning context given by the previous model
Probably we should strip the reasoning parts after the message completes even if no model switch is made imo
Instead only the actual final message should be kept