Source

Transcript

Explain it to me like I’m 80

Mom: What’s with these AIs escaping and hacking people? That sounds scary.

Me: You park your car at the top of a hill. You leave a note on the dashboard that says “please be good” and then you pull the parking brake. The car careens down the hill, smashes everything up, and you put out a press release saying “My car ignored my instructions! It hallucinated! It went rogue!” Your insurance company not only believes you, but invests a hundred million dollars in your company.

Mom: Jesus Fucking Christ

  • rumba@lemmy.zip
    link
    fedilink
    English
    arrow-up
    1
    ·
    18 hours ago

    I didn’t downvote, but at the same time, I don’t for a moment think that’s what happened. They purposefully trained them with exploit knowledge, set them loose in a weakly secured sandbox with the goal of finding exploits, and let them run without intervention.

    Just adding an agent to a computer does nothing; something needs to trigger it. If i create an agent with access to a kali or backtrack box and turn all guardrails off, I still have to tell it to start looking for exploits.

    • TrickDacy@lemmy.worldM
      link
      fedilink
      arrow-up
      2
      arrow-down
      1
      ·
      15 hours ago

      Words are squishy. OP getting a negative reaction to me is silly because it takes multiple assumptions to take that negatively.

      I took their words to mean “LLMs can go off the rails and do things you didn’t ask”. This is 100% true. The other day I asked an LLM to review my PR. It ended up committing code to fix a supposed issue my coworker commented about. At no point did anyone ask for anything but a review, never for it to make changes.

      I got no impression that OP thinks LLMs are running wild on their own now. I just got that they think they are unpredictable. Which is true.