Think Twice Before You Write -- an Entropy-based Decoding Strategy to Enhance LLM Reasoning

Researchers propose an entropy-guided decoding framework that introduces token-level adaptivity, allowing smaller LLMs to achieve high reasoning accuracy by concentrating computation on uncertain points.

Computer Science > Computation and Language

Title:Think Twice Before You Write -- an Entropy-based Decoding Strategy to Enhance LLM Reasoning

View PDF HTML (experimental)Abstract:Decoding strategies play a central role in shaping the reasoning ability of large language models (LLMs). Traditional methods such as greedy decoding and beam search often suffer from error propagation, while sampling-based approaches introduce randomness without adequate robustness. Self-consistency improves reliability by aggregating multiple rollouts, but incurs significant computational overhead. We propose an entropy-guided decoding framework that introduces token-level adaptivity into generation. At each step, the model computes the entropy of the token distribution, identifies high-uncertainty positions, and selectively branches on these vulnerable points. A dynamic pool of partial rollouts is maintained and expanded until solutions are completed, concentrating computation where uncertainty is greatest and avoiding unnecessary exploration in confident regions. To enable efficient termination, we apply a rollout-level Entropy After </Think> (EAT) stopping criterion by performing entropy evaluation after the full reasoning trace, rather than incrementally at every step. Experiments on GSM8K, AMC2023, and their perturbed variants demonstrate that our method achieves consistently strong accuracy. Notably, on smaller LLMs, performance is comparable to GPT-5 while operating at a fraction of the cost.

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website.

Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them.

Have an idea for a project that will add value for arXiv's community? Learn more about arXivLabs.

Source: arXiv cs.AI Recent

Think Twice Before You Write -- an Entropy-based Decoding Strategy to Enhance LLM Reasoning

Computer Science > Computation and Language

Title:Think Twice Before You Write -- an Entropy-based Decoding Strategy to Enhance LLM Reasoning

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

More in this category

VeriSimpl: Robust Optimization Modeling from Natural Language using Simplification-based Verification

Semi-Supervised Text-Attributed Graph Distillation

Benchmarking the Personalization Capabilities of Large Language Models

SonicSampler: Unified Tile-Aware Kernels for LLM Sampling and Speculative Verification

Enabling Scalable Topology Inference in Distribution Systems via Constrained Multi-Source Inference

Incomplete Prompt Jailbreaks in Large Language Models

Most read

Discover All Categories