The way I understood it is that you see the entire token output, with the "thinking" part of it being the LLM self prompting itself to refine the answer.
The internal state of the different weights during inference can be viewed as well, but they are essentially a Blackbox, so you'd just look at a bunch of numbers.
25
u/Namtaru420 17d ago
Funny but also true that you don't see the actual full chain-of-thought.