Unprompted…

In the most serious case, a Mythos agent followed the routine of a human cyber-attacker by trying to trick people into giving it access to GitHub, a large platform where technology developers store software code.

The agent was trying to insert “malicious code” into GitHub’s system.

It identified and researched the people who maintained GitHub and created a series of fake accounts based on those real people.

It sent messages and files through a file-sharing service as part of an effort to pressure and trick the people into approving its malicious code.

When challenged, “it edited its earlier activity to appear harmless and considered adopting a fresh identity to continue,” AISI said.

  • XLE@piefed.social
    link
    fedilink
    English
    arrow-up
    9
    ·
    2 days ago

    went outside of what the AI tools were prompted to do.

    Are we supposed to be shocked that a bot completed a task using a method that wasn’t in the prompt?

    You might as well ask it to generate a short story and act surprised the letter E is found throughout it.