The chain of thought is super useful in so many ways, helping me: (1) learn, way beyond the final answer itself, (2) refine my prompt, whether factually or stylistically, (3) understand or determine my confidence in the answer.
It uses them as tokens to direct the chain of thought, and it is pretty interesting that it uses just those works specifically. Remember that this behavior was not hard-coded to the system.
What do you mean?
I was referring to just the chain of thought you see when the "DeepThink (R1)" button is enabled.
As someone who LOVES learning (as many of you too), R1 chain of thought is an infinite candy store.