Sandboxed experiment found itself a zero day, escaped onto the open internet and validated scary predictions about rogue agents

  • hayvan@piefed.world
    link
    fedilink
    English
    arrow-up
    19
    ·
    edit-2
    11 days ago

    Take these admissions with truckloads of salt. While agents finding vulnerabilities to vioate their containment isn’t impossible, “our new product is too dangerous, trust me bro” declarations have proven to be great PR, so they have massive incentives to exaggerate whatever happened.

    • Rothe@piefed.social
      link
      fedilink
      English
      arrow-up
      9
      ·
      11 days ago

      Yeah, the Anthropic PR method. They are all a bunch of lying scumbags desperately trying to prop up the line as they are hemorrhaging capital.

    • Sundray@lemmus.org
      link
      fedilink
      English
      arrow-up
      8
      ·
      11 days ago

      OpenAI: Whoa, did you guys see that? Man, it’s almost like a government might want to give us billions of dollars for something like that, huh?