Humans essentially do "next token prediction" too - there's always a choice between the next actions to take and they pick a good one based on what has happened in the past.
That doesn't really limit how clever we can get internally when picking the next action.
That doesn't really limit how clever we can get internally when picking the next action.