Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

> If the intermediate tokens represent reasoning or thought, you would expect "aha" to occur after the thoughts that led to the realisation, including the thoughts encoding the explanation: they don't have any other state.

Yes they do, they have their KV caches-- it's a pure function of the input tokens, sure but that doesn't prevent it from containing latent 'insight'. LLMs can and do pre-form the tokens they're expecting to output multiple steps in the future.

I wouldn't argue that the 'aha' means anything, but the structural argument that it can't that I think you're making isn't sound.



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: