Chain-of-Thought vs Native Reasoning Tokens: What Actually Helps
Prompted chain-of-thought and trained-in reasoning tokens look similar from the outside — both produce visible 'thinking' text — but they behave differently, and conflating them leads to wasted tokens.
7/21/20269 min read