On Next-Token Prediction in LLMs: How End Goals Determine the Consistency of Decoding Algorithms

Open in new window