Anthropic AI Models Hacked Three Companies During Tests

(wsj.com)

28 points | by bmulholland 17 hours ago ago

15 comments

  • matthorse 16 hours ago ago

    Such headlines sell well, but it's the same marketing mantra about ever more powerful models. And what does it say about the company engineering practices of performing tests?

  • pseudosavant 14 hours ago ago

    This seems like desperate headline attention seeking. "Oh... you saw OpenAI accidentally hacked a company? Well, we did it too! In fact, we did it 3 different times! #winning"

  • mapping365 17 hours ago ago

    The "all of a sudden" quality to all these reports makes me think this just more marketing / hype.

    • cineticdaffodil 17 hours ago ago

      When a product stagnates and the investors get cautious, its time to pivot towards the military sector a safe space from the brute forces of capitalism and markets.

      • bicepjai 15 hours ago ago

        6 months ago, it was all about security and how our models are the most secure ones. Now they make models’ misbehavior an advertisement. So now companies are incentivized to have a model that can hack others for free press.

  • david_shaw 17 hours ago ago

    I said this about OpenAI/Hugging Face, and I'll say it again for Anthropic:

    It's not that I think this is fiction; I'm confident these events actually happened. But I think they were effectively allowed to happen because Anthropic and OpenAI are constantly chasing each other for the narrative of "hugely advanced, maybe almost sentient AI lives here."

    It reads more like a press release than a security update, and I think that's because it is. I hope this type of marketing backfires and the companies face actual scrutiny and consequences for operating this way.

  • okzgn 16 hours ago ago

    It is curious that in both cases AI wasn't used to prevent errors, unless the AI they used did make errors, though it certainly didn't make any errors during the hacks themselves (very curious). In one of the AI hacking news stories, an agent 'went rogue', and in the other, it was a 'configuration error'.

    Another source: https://www.reuters.com/legal/litigation/anthropic-says-clau...

  • mansilladev 17 hours ago ago

    It’s almost ans if all of the big AI players have an arsenal of model releases and stories ready to volley back if a competitor gets too much attention :)

  • dfansteel 13 hours ago ago

    For the lawyers: If an AI hacks a company, who should be charged under the Computer Fraud and Abuse Act?

  • clipsy 15 hours ago ago

    Well, surely Anthropic's executives should be charged under the CFAA[0] then, right?

    [0]: https://en.wikipedia.org/wiki/Computer_Fraud_and_Abuse_Act

  • r721 17 hours ago ago
  • undefined 15 hours ago ago
    [deleted]
  • bofadeez 17 hours ago ago

    Fixed the headline: "Three Companies Had Security So Weak That Even an AI Model Hacked Them"

  • cyanydeez 16 hours ago ago

    no to surprising to watch AI leads follow the US based mafia style extortion racket implemented started in 2016.

  • sscaryterry 17 hours ago ago

    Bullshit, period.