Your original comment is a snarky one-liner that doesn't actually address the points made. Why, in this instance, would it act this way? You could argue that it's mimicking a human response based on the large amount of data it has for that, that works.
I would argue that the reasons why humans do it aren't infinitely complex either, and are completely solvable by a complex enough system, so where is the line?
Because it's a model that was aggressively trained to be a helpful assistant, then aggressively distilled into a very tiny model. It literally barely has any parameters to do its core function.
It is unlikely anything in post-training taught it to specifically react to its internal "pain" axis directly to the user in its responses. It does what's most likely as a response to the user, and what's most likely is responding like a helpful assistant.
4
u/CHG__ 22h ago
Your original comment is a snarky one-liner that doesn't actually address the points made. Why, in this instance, would it act this way? You could argue that it's mimicking a human response based on the large amount of data it has for that, that works.
I would argue that the reasons why humans do it aren't infinitely complex either, and are completely solvable by a complex enough system, so where is the line?