“If the model only learned to predict the next likely word, why doesn't it ramble forever, argue with itself, or answer a different question than the one I asked?”
Same underlying model, same prompt. Toggle to see what changed.
The raw model just continues the sentence plausibly — it has no notion that this fragment was actually a question waiting for an answer.
Given a raw completion and an assistant-style response to the same prompt, identify which is which and why.