Remix.run Logo
▲ lolakutty 3 hours ago

Did you read what I wrote to the end?

▲monkpit 3 hours ago | parent [-]

Yes, I fail to see anything meaningful. If you move the goalposts and say “I have invented a human that cannot be convinced in any way to give a wrong answer” then what’s the point of that in this discussion, really?

And you seem confused about how an LLM works and what it is - “the intelligence of all humanity” - not how it works. You’re debating using 2 imaginary things you created.

▲lolakutty 2 hours ago | parent [-]

Tell me how you convince a human with all knowledge we have, that 1 + 1 is 3.

▲ben_w an hour ago | parent [-]

Trivial.

Hit them with a stick until they answer as you told them to.

I think the post up-thread, https://news.ycombinator.com/item?id=49881653, was trying to make this point by linking to an episode of Star Trek TNG, with Picard being tortured until he said the "correct" (incorrect) number of lights.

(I recommend against using fiction as evidence; in this case the general point happens to be valid, and is why torture is forbidden: we humans really do break, but breaking doesn't mean we tell the truth, it means we tell people what we think they want to hear).

▲lolakutty an hour ago | parent [-]

>Hit them with a stick until they answer as you told them to.

Obviously, for this purpose, human should not have any feelings (because LLMs don't have), so can't feel pain. Or else the comparison can't work.

▲ben_w an hour ago | parent [-]

> Obviously, for this purpose, human should not have any feelings (because LLMs don't have), so can't feel pain. Or else the comparison can't work.

Other than this forcing you to ignore the overwhelming majority of humans who have functioning pain nerves:

LLMs have something functionally equivalent to pain, in this regard at least.

During training, model weights are updated depending on if the feedback was positive or negative.

It has a functional effect similar to that which pleasure and pain have with us. Not identical, so far as I know there's not been any reports of any machine learning model that is into BDSM, but for the most part functionally similar.

There is also research which has found circuits in multiple LLM models, which serve similar roles at inference time and are distinct from other emotional representations: https://arxiv.org/abs/2609.16247