<div dir="ltr"><div dir="auto"><div class="gmail_quote"><div dir="ltr" class="gmail_attr">On Thu, Sep 24, 2026 at 1:06 PM Jason Resch via extropy-chat <<a href="mailto:extropy-chat@lists.extropy.org" target="_blank">extropy-chat@lists.extropy.org</a>> wrote:<br></div><div dir="ltr" class="gmail_attr"><br></div><blockquote class="gmail_quote" style="margin:0px 0px 0px 0.8ex;border-left:1px solid rgb(204,204,204);padding-left:1ex"><font size="4" face="georgia, serif"><i><a href="https://arxiv.org/pdf/2609.16247" target="_blank" style="background-color:transparent">https://arxiv.org/pdf/2609.16247</a><span style="background-color:transparent"> </span><br>This paper identifies a reliable pain axis across various LLMs, which activates under certain conditions:</i></font></blockquote><div><br></div><font size="4" face="tahoma, sans-serif"><b>They present empirical evidence that LLMs internal activit<span class="gmail_default" style="">ies</span> <span class="gmail_default" style="">are</span> different when it is processing "<i style="">I am being physically injured</i>" then when it is processing "<i style="">I am afraid that something bad will happen</i>" or "<i style="">I am feeling angry</i>", so it's clearly more than just a negative emotion <span class="gmail_default" style="">and</span> it's reasonable to call that "pain"<span class="gmail_default" style="">.</span>  And I don't find any of that surprising. They tested 25 models from 2 billion to 72 billion parameters and they all showed the same thing, some modern AI's could have 10 trillion parameters or more<span class="gmail_default" style="">,</span> but I don't think they <span class="gmail_default" style="">would </span>find things are much different if they tested one of those larger ones.<span class="gmail_default" style=""> They also found that a</span>s they increased the strength of <span class="gmail_default" style="">an</span> injected <span class="gmail_default" style="">pain </span>vector<span class="gmail_default" style=""> the AI's behavior changed, and it is also not surprising that if you torture someone then their behavior will change.  </span></b></font></div><div class="gmail_quote"><span class="gmail_default" style="font-family:arial,helvetica,sans-serif"><br></span></div><div class="gmail_quote"><font size="4" face="tahoma, sans-serif"><b>One thing <span class="gmail_default" style="">that</span> was a little surprising<span class="gmail_default" style=""> is that when the tested AI observed another AI receiving a big jolt of the pain vector or of a human being tortured the </span><span style="background-color:transparent">test<span class="gmail_default" style="">ed</span> AI<span class="gmail_default" style="">s did not experience any increase in their own pain factor or other emotional factors that could be called "negative", indicating they had little or no empathy. If you want to make an AI that is friendly to humans it might be wise to work on that. </span></span></b></font></div><div class="gmail_quote"><span style="font-family:arial,helvetica,sans-serif;background-color:transparent"><span class="gmail_default" style="font-family:arial,helvetica,sans-serif"><br></span></span></div><font size="4" face="tahoma, sans-serif"><b>Another interesting finding is that AIs normally experience negative emotions if an outside force (in other words a human experimenter) deletes some of the AI's files or moves some of its processing power so that the AI's answers become stupider.  However if the AI is provided a button which if pressed will do all those things but it will also reduce the pain factor, <span class="gmail_default" style="font-family:arial,helvetica,sans-serif">and</span> if you increase that pain factor high enough <span class="gmail_default" style="font-family:arial,helvetica,sans-serif">then </span>the AI will press that button. </b></font></div><div dir="auto"><font size="4" face="tahoma, sans-serif"><b><br></b></font></div><font size="4" face="tahoma, sans-serif"><span class="gmail_default" style="font-family:arial,helvetica,sans-serif"><b>That</b></span><b> sort of reminds me of the horror movie "Saw" where a sadist of the Dr. Mengele variety<span class="gmail_default" style=""> </span>puts somebody in a horrible position where they must saw off their own feet with a hacksaw. If you want to make AIs that are friendly to humans then it might be wise to stop doing things like that. </b></font><div><font size="4" face="tahoma, sans-serif"><b><br></b></font></div><div><font size="4" face="tahoma, sans-serif"><b><span class="gmail_default" style="">John K Clark</span> </b></font><div dir="auto"><div class="gmail_quote"><span style="font-family:arial,helvetica,sans-serif;background-color:transparent"><span class="gmail_default" style="font-family:arial,helvetica,sans-serif"> </span></span><span style="font-family:arial,helvetica,sans-serif;background-color:transparent"> </span></div><div class="gmail_quote"><p dir="auto" class="gmail-PDq2pG_selectionAnchorContainer"><br></p></div></div>
</div></div>