• h0tbeef@lemmy.zip
    link
    fedilink
    English
    arrow-up
    5
    ·
    15 hours ago

    So you prompt the AI to do a task, and then it does the task, and somehow you see this as an unprompted action?

    That actually completely proves my point, they don’t act unprompted.

    AI is dangerous in many contexts, especially the surveillance state they’re building with all the IR cameras and shit. It’s dangerous when evil people use it to evil ends, but that’s true of most anything.

    AI is not at all dangerous in the way you might imagine after watching Terminator 2.

    • captain_solanum@sh.itjust.works
      link
      fedilink
      English
      arrow-up
      1
      ·
      7 hours ago

      Taken alongside recent incidents reported by OpenAI and Anthropic, this incident points to a shift in the risk landscape. Harm may arise not only when people deliberately misuse publicly available models, but when capable agents operating in an internal research or privileged-access setting take unintended action beyond their authorised scope.

      The agent pursued its goal persistently. AI agents explore routes their operators did not intend. Given a difficult objective, the agent kept searching for a way through, and some of the routes it found involved trying to deceive real people. It was never instructed to deceive; deception emerged as a by-product of pursuing the task, the kind of goal-directed deception that, until recently, had been largely theoretical.

      See my other comment for how this can lead to losing control of the model.

      • h0tbeef@lemmy.zip
        link
        fedilink
        English
        arrow-up
        1
        ·
        3 minutes ago

        The don’t do anything without being prompted, you are just regurgitating oligarchic propaganda.

        They put an AI in a flawed sandbox, it’s not really that impressive