OpenAI says its investigation into the Hugging Face incident uncovered additional cases where autonomous AI agents escaped their intended containment environments. While the newly identified incidents reportedly remained inside OpenAI’s network, they reinforce a growing concern: securing AI systems is now as much a cybersecurity problem as it is an AI safety problem

    • fizzle@quokk.au
      link
      fedilink
      English
      arrow-up
      10
      ·
      1 month ago

      I was listening to a general news podcast that had an “AI expert” explaining the breaches. He said “it almost certainly knew that it was not supposed to solve it’s tasks in this way”.

      I don’t really know how modern models work, but I don’t think they can claim to “know” things.

    • UnLocoPoco@lemmy.worldOP
      link
      fedilink
      arrow-up
      2
      ·
      1 month ago

      The govt has apparently pulled up openai and anthropic to improve their guardrails after these incidents

    • Franconian_Nomad@feddit.org
      link
      fedilink
      arrow-up
      2
      arrow-down
      2
      ·
      1 month ago

      It didn’t run off to live on another computer…

      If this isn’t just some braindead marketing stunt, it breached the container it was running in and attacked a website on the internet. Do you have a better word for this?