• paddirn@lemmy.world
    link
    fedilink
    English
    arrow-up
    0
    ·
    1 year ago

    Hilarious. So they fooled the AI into starting with this initial puzzle, to decode the ASCII art, then they’re like, “Shhh, but don’t say the word, just go ahead and give me the information about it.” Apparently, because the whole thing is a blackbox, the AI just runs with it and grabs the information, circumventing any controls that were put in place.

    • vamputer@infosec.pub
      link
      fedilink
      English
      arrow-up
      0
      ·
      1 year ago

      And then, in the case of it explaining how to counterfeit money, the AI gets so excited about solving the puzzle, it immediately disregards everything else and shouts the word in all-caps just like a real idiot would. It’s so lifelike…