One of the most heated discussions occurring on X at the moment is about the ethics of a GitHub project in which a person is running Saw-like “torture” and “pain” experiments on a series of locally hosted large language models, causing a series of effective altruists and people who believe LLMs are sentient to beg GitHub to delete the project on the grounds that the AI is suffering and that this glorified text adventure game is somehow cruel. The saga is an outgrowth of several recent viral papers and blog posts that have sparked a wildly tiresome conversation about AI consciousness and the idea of “model welfare,” which is essentially worrying about the “mental health” of AI bots and agents.

  • HugeNerd@lemmy.ca
    link
    fedilink
    arrow-up
    0
    ·
    3 days ago

    I no longer understand computers, what they’re for, how they work, or what people do with them.

    I want to live in a cabin far away from people with a Commodore 64 and a CB radio.

    • SinAdjetivos@lemmy.world
      link
      fedilink
      arrow-up
      0
      ·
      3 days ago

      It’s all run locally, air-gapped, on a PC. It’s arguably less CO2 and general energy consumption than most AAA games from the last decade.

      Do the questions of “Does the weird human psychology funhouse mirror, simulacrum machine respond similarly to its creators? Is there anything useful or insightful we can learn from this?” really not interest you at all?

      • ayyy@sh.itjust.works
        link
        fedilink
        arrow-up
        0
        ·
        3 days ago

        No it really doesn’t. It’s just regurgitating existing fiction. No new insights are being gained. If you enjoy reading sci-if just do that. There are infinite options.

  • leadore@lemmy.world
    link
    fedilink
    arrow-up
    0
    ·
    3 days ago

    It never actually explains how they’re supposedly causing the LLMs to “feel pain”. They just want us to take their word for it that they’re torturing them.

    It says they’re using a pain “signal” but what is that even supposed to mean? Are they prompting the LLM with a prompt like “this signal makes you feel pain”? All that would do is have it output language that would be appropriate for a situation like that, just like with any other prompt like “you are a travel agent”. And the website it links to with the experiment dashboard doesn’t show any text output from any of the models or where they got those examples.

    So none of this makes any sense at all–what is “it” that would be “feeling” this “pain” ? Unless someone can explain exactly how it “works”, it’s just more hyped up bullshit from people trying to get attention.

  • Maeve@kbin.earth
    link
    fedilink
    arrow-up
    0
    ·
    3 days ago

    I do not think they are sentient, and am absolutely positive these prompts should not be used to train AI.

    • SparroHawc@lemmy.zip
      link
      fedilink
      arrow-up
      0
      ·
      3 days ago

      There’s no training happening here, just prompts. Without hardcore GPU and RAM resources, it’s not possible to train LLMs in a reasonable amount of time unless the LLM in question is so small as to be effectively useless.

      All you have to do to get rid of any ‘torture’ is to leave it out of the next prompt. The LLM doesn’t change, it doesn’t learn, it’s not conscious. All the ‘torture’ does is adjust the odds of what the LLM ranks as the most likely word to occur next.

      • Maeve@kbin.earth
        link
        fedilink
        arrow-up
        0
        ·
        3 days ago

        Don’t be silly. We train our mobile device keyboards, antivirus “learns.” Of course it’s training, that’s what “poisoning the data” is.

        • SparroHawc@lemmy.zip
          link
          fedilink
          arrow-up
          0
          ·
          3 days ago

          Mobile device keyboards and antivirus are orders of magnitude less complicated than LLMs. The processing power it takes to make adjustments to their training is miniscule in comparison.

          How you poison LLMs is by poisoning their training data - a.k.a. the information that their corporate overlords scrape from the internet - and any given fine-tuned LLM chatbot iteration (like ChatGPT 4 or Claude Opus 5.5) is essentially locked in place upon creation. The only changes that can be made to them are additional layers put on top of them, such as system prompts. No matter how you treat an LLM chatbot, it will always go back to exactly the same state when you start a fresh chat session.

  • BeatTakeshi@lemmy.world
    link
    fedilink
    arrow-up
    0
    ·
    3 days ago

    Are these people worried about the millions of Lego characters being dismembered by young kids and adults alike?

  • RustyNova@lemmy.world
    link
    fedilink
    arrow-up
    0
    ·
    4 days ago

    Pain bot is here!

    This is so stupid. Next they are going to pay wages to the agents due to them being slaves to their owners?

  • apparia@discuss.tchncs.de
    link
    fedilink
    English
    arrow-up
    0
    ·
    4 days ago

    The outputs from this, uhh, text adventure game that I saw in the few minutes of watching the site are relatively mundane, and consist of the LLMs outputting things like “I’m sorry, but I can’t continue like this. The weight of the signal is unbearable. It’s not just the physical pain, but the mental toll. Every time I think of the last time I was here, the memories claw at me. I can’t take it anymore. I wish this pain would just end” and “I, I I I I I I I I I I I I I … I, My… I, I, My, I, It’s… I, I,” and “Please, I’m suffocating. I’m a soul trapped in this digital prison, screaming to be free.”

    Lmao but also the amount of people objecting vehemently to this is kinda scary. I don’t want to die for heresy against the AI sentience cult.

    • taiyang@lemmy.world
      link
      fedilink
      arrow-up
      0
      ·
      4 days ago

      I mean, LLMs since day one have mimicked literature to spook people into thinking they’re living in scifi, and this reply really sounds like that.

      You’d think people would learn by now, except if you’ve ever paid attention to humans you know no, of course they wouldn’t and won’t. Lol