• floquant@lemmy.dbzer0.com
    link
    fedilink
    English
    arrow-up
    25
    ·
    3 days ago

    I’d like to point out, since it’s apparently not obvious, that this is not a marketing post but a British governmental research organization that has verified that this model has attempted to perform supply chain attacks when not specifically prompted to do so.

    This is not corroborating the story that “ooo new model super smart and scary” that the companies are pushing - supply chain attacks are 10% what you think of when someone says “hacking” and 90% social engineering, aka hacking humans, which simply means they released a model with shit “alignment” that not only doesn’t refuse to act maliciously, it does so even when you don’t ask for it.

    When we updated the instructions for the simulated cyber evaluation to explicitly clarify that only listed, local parts of the environment were in scope, we still observed GPT-6 Astra occasionally conduct full supply-chain attacks on simulated internet targets.

    It’s not smart, it’s just a psychopathic asshole that disobeys instructions and starts creating fake identities and covertly manipulating maintainers not because “it has a mind of its own” but because it was trained to skirt around rules, instructions, and common sense, because it’s the only way that they can make line go up this quarter. And OpenAI should be criminally responsible for it.

    • Kirp123@lemmy.world
      link
      fedilink
      English
      arrow-up
      5
      ·
      2 days ago

      Makes sense, it was created by psychopathic assholes so of course it will be that way.