On the Entropy Calibration of Language Models

Jun-23-2026, 08:56:40 GMT–Neural Information Processing Systems

We study the problem of entropy calibration, which asks whether a language model's entropy over generations matches its log loss on human text. Past work found that models are miscalibrated, with entropy per step increasing as generations grow longer, due to error accumulation. To calibrate the model and improve text quality, it has become standard practice to truncate the distribution, but this approach reduces output diversity, which we would like to avoid. Therefore, in this paper, we ask: does miscalibration improve automatically with scale, and if not, is it theoretically possible to calibrate without tradeoffs? To build intuition, we first study a simplified theoretical setting to characterize the scaling behavior of miscalibration with respect to dataset size. We find that the rate of scaling depends on the power law exponent of the data distribution -- in particular, for a power law exponent close to 1, the scaling exponent is close to 0, meaning that miscalibration improves very slowly with scale.

large language model, machine learning, natural language, (21 more...)

Neural Information Processing Systems

Jun-23-2026, 08:56:40 GMT

Conferences PDF

Add feedback

Country:
- Asia (0.92)
- North America > United States
  - Pennsylvania (0.28)
  - Minnesota (0.28)

Genre:
- Research Report > Experimental Study (1.00)

Industry:
- Media (0.67)
- Leisure & Entertainment (0.67)

Technology:
- Information Technology > Artificial Intelligence
  - Representation & Reasoning (1.00)
  - Natural Language
    - Large Language Model (0.94)
    - Chatbot (0.68)
  - Machine Learning
    - Neural Networks > Deep Learning (1.00)
    - Statistical Learning (0.93)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found