Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

While I tend to agree on the overall sentiment, I think this rebuke is inaccurate. Some of these "reasoning" models are trained using "Chain-of-Thought" where the model is presented explicit, intermediate reasoning steps (either by a human or some automation) that supposedly get it closer to the correct answer. These intermediate steps are what was originally called "thinking traces" - not what the model produces to mimic them.

But yes, anthropomorphizing model outputs leads to worse outcomes.



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: