On the Token Value Inequality in Efficient Reasoning
Chain-of-Thought reasoning has enabled large language models to achieve substantial performance gains on complex tasks. However, these gains come at the cost of dramatically increased token consumption. This raises a fundamental question: is every token in the reasoning trace equally valuable? We present a diagnostic and optimization framework grounded in a key empirical finding: the value of…
We haven't written up this one. arXiv cs.AI has the full story — the link below goes straight to it.