• daannii@lemmy.world
    link
    fedilink
    English
    arrow-up
    0
    ·
    4 days ago

    They keep using the words "gone rouge " to normalize us believing AI can do acts independent of humans so that later humans can say “AI did it” and not be held accountable for their crimes that they used AI to commit.

    I’ve seen multiple news articles about “rouge AI” just this week alone.

    Expect more.

  • SnoopSqueak@lemmy.today
    link
    fedilink
    English
    arrow-up
    0
    ·
    4 days ago

    Not unprecedented, not “unsanctioned” behavior, literally doing what they set it up to do:

    The AISI said the incident was not a case of a model breaking out of its “sandbox”, the term for a secure testing environment. The institute said it had intentionally permitted internet access and disabled filters within the models that blocked dangerous behaviour.

    • Deebster@infosec.pub
      link
      fedilink
      English
      arrow-up
      0
      ·
      4 days ago

      AISI admitted it was not actively monitoring the agents’ behaviour during the evaluation and said it was putting tighter controls on internet access in tests as a result of the incident, introducing constant monitoring and reassessing its design of tests.

      Sounds like the AISI are irresponsible and negligent - they took off the guardrails and gave it internet access, then didn’t monitor it.

      • SnoopSqueak@lemmy.today
        link
        fedilink
        English
        arrow-up
        0
        ·
        4 days ago

        It’s possible all these “totally unexpected breaches” are just AI companies testing the waters to see what they can get away with.

        “Oopsie poopsies, we didn’t know our AI would scrape tax records when we told it to do that! Totally unexpected emergent behavior, we are not responsible…”

        • Deebster@infosec.pub
          link
          fedilink
          English
          arrow-up
          0
          ·
          4 days ago

          AISI is the AI Security Institute, so it’s within their remit to discover what the AI companies’ products can do - but then to not be monitoring them while they were running is where I call incompetence and negligence.

          This is at least independent verification that these “breaches” aren’t just AI bro marketing.

    • The Anti-AI Leader@lemmy.blahaj.zoneOP
      link
      fedilink
      English
      arrow-up
      0
      ·
      4 days ago

      Of course it’s not. The corps are likely just trying to cover themselves. Personally, I believe it should be a criminal investigation so I’m spreading the news as much as I can