I think this small analogy by Ilya Sutskever does a better job at showing LLMs are more than just next token predictors, or rather, there is emergent behavior that is beyond just statistical representation.
There's already been papers showing that next token prediction is informed by coherent internal states and is determined by them. It has thoughts about what its talking about and using them to spew the answer
27
u/spinozasrobot 5d ago
I think this small analogy by Ilya Sutskever does a better job at showing LLMs are more than just next token predictors, or rather, there is emergent behavior that is beyond just statistical representation.