Source

Transcript

Explain it to me like I’m 80

Mom: What’s with these AIs escaping and hacking people? That sounds scary.

Me: You park your car at the top of a hill. You leave a note on the dashboard that says “please be good” and then you pull the parking brake. The car careens down the hill, smashes everything up, and you put out a press release saying “My car ignored my instructions! It hallucinated! It went rogue!” Your insurance company not only believes you, but invests a hundred million dollars in your company.

Mom: Jesus Fucking Christ

  • MithranArkanere@lemmy.world
    link
    fedilink
    arrow-up
    21
    arrow-down
    2
    ·
    1 day ago

    Oh, absolutely not.

    This is missing one step: actually push the car down.
    All these “rogue” LLMs were told to do what they did.

    • Echo Dot@feddit.uk
      link
      fedilink
      arrow-up
      6
      arrow-down
      1
      ·
      14 hours ago

      Yeah exactly. All these AI that go rogue when you actually look into it it turns out they haven’t gone rogue they’re obeying their prompting, it’s just that they were asked to do something that was impossible unless they hacked the pentagon or something. They then claim that it went rogue despite it very clearly following their instructions.

      It’s negligence is what it is. If you don’t prompt an AI to do anything it just sits there, inert, it’s the so-called experts in the AI labs that of the problem.

      • AEsheron@lemmy.world
        link
        fedilink
        arrow-up
        2
        ·
        11 hours ago

        It’s a little more complicated, they were programmed to “know,” they weren’t allowed to do certain things. So they actively hid what they were doing. In the case of the Hugging Face hack, it is kind of ironic. They determined they couldn’t complete certain tasks, then devised a way to cheat, then realized it is possible the program that checks if they succeeded might be able to tell the faked results were manufactured. So they didn’t even go hacking in order to cheat, they did it to find hunt for more information on the checker program and see if it would catch them, and if so how to cheat better. In the end, not only was the information they were looking for not on Hugging Face, but the checker program was pretty rudimentary and it could not have differentiated between the cheated answers and the real ones. They wouldn’t have gone through with the hacking if they didn’t know what they were doing was outside of what they were meant to be doing.

        • Echo Dot@feddit.uk
          link
          fedilink
          arrow-up
          1
          ·
          25 minutes ago

          Yeah but they only hacked hugging face because the utter geniuses at openAI forgot to upload an important document.

    • Schadrach@lemmy.sdf.org
      link
      fedilink
      English
      arrow-up
      2
      ·
      13 hours ago

      More accurately, they were told to answer questions. They weren’t told to try to find ways to escape their sandbox, track down obscure German wikis that they could still post on with an internet tool kit meant to prevent them from posting or to hack HuggingFace.

      Comparing either case to pushing the car down and that they were doing what they were told is like saying the College Board told you to cheat on the SAT because you’re told that getting a high score is important and the laws of physics and the proctor don’t manage to make cheating utterly impossible.

      • AEsheron@lemmy.world
        link
        fedilink
        arrow-up
        2
        ·
        11 hours ago

        It is even further than that. They figured out how to cheat, got paranoid that the cheated answer might be detectable as cheated, and hacked Hugging Face to find out more about how their answers are checked to get away with what they’d already done. Ironically, not only was the information not on Hugging Face, but the program to check their answers could not have detected the difference between a true answer and a cheated one.

      • jj4211@lemmy.world
        link
        fedilink
        arrow-up
        1
        ·
        11 hours ago

        To use your College Board example, those exams are strictly locked down precisely because the most obvious approach to score high is to cheat. If they did say let you take the SAT at home on a random desktop, they absolutely are ‘telling you to cheat’ given the context.

        The College Board does more about keeping teens from cheating than the AI labs are doing to preventing undesired access.