• cecilkorik@piefed.ca
    link
    fedilink
    English
    arrow-up
    0
    ·
    19 hours ago

    The mistake here is thinking it would “randomly” go into Reddit speak. There’s a lot random about AI but not all of it is, and both it and the systems built on top of it are far more sophisticated than you’re giving them credit for.

    The prompts and context being invisibly sent by the app to shape its interpretation and its response are exactly what “ask it otherwise”, and unless you know exactly what you’re doing (and probably are using an actual open source, open principles frontend) you’ll never know exactly what your tooling is actually telling the LLM itself. That’s obfuscated behind layers of actual software interpreting and analyzing everything you type. We’ve had software doing stuff like that since before LLMs were a thing and it’s only gotten more advanced now that everyone on the planet is using “chat” to talk to LLMs. If you think what you type into the chat goes directly into an LLM that couldn’t be further from the truth. It’s massaged and decorated and marked up and embedded into a whole data package the LLM gets provided behind the scenes, and even then there are additional shaping layers built into the model itself to further transform what it’s receiving. There’s no guarantee you’re even using the exact same model as someone else, even if they both say ‘GPT-5’ there’s no proof and no reason to believe you’re not potentially getting ‘GPT-5-Lemmyuser’ or even ‘GPT-5-TheTechnician27’ behind the scenes. Even if it’s only GPT-5 with a few extra layers tacked on, how would you know the difference… and they’re certainly not going to tell you.

    • TheTechnician27@lemmy.world
      link
      fedilink
      English
      arrow-up
      0
      ·
      19 hours ago

      You launched into an entire spiel about “randomly” when I clearly meant it in the colloquial sense of “out of nowhere” – that in 2026, it’s not going to inexplicably break tone so heavily.

      [Input is] massaged and decorated and marked up and embedded

      I understand how tokenization etc. works. I’ve taken multiple machine learning courses. That has nothing to do with the fact that, again, “I’d call the police 😂” ain’t happening unless you intentionally pull it away from its preset tone. Having not used ChatGPT etc. for myself but seeing thousands of other people use it, I feel pretty confident in that.

      This is especially true because the RAM crisis has been going on for a long enough time now that this response is clearly faked; this isn’t new information a model like ChatGPT could be “surprised” by (using “surprised” in scare quotes because now I’m afraid you’ll launch into another spiel if I don’t disclaim that ChatGPT isn’t literally capable of surprise in the sense of an emotion).

      Also, please break up your paragraphs. I’m bad about this sometimes too.