"Next-token predictor" is the wrong (or at least incomplete) mental model for post-trained LLMs. I hadn't thought of it like that before. It's always good to understand the tools you're using. gmcgoldr.github.io/2026/09/04/llm-n…
gmcgoldr.github.io
Stop Thinking of LLMs as Next-Token Predictors
Stop Thinking of LLMs as Next-Token Predictors