• Zedstrian@sopuli.xyz
    link
    fedilink
    English
    arrow-up
    40
    ·
    edit-2
    7 days ago

    The fact that so many people don’t see these ‘hacking’ stories for the pure and unabashed marketing that they are is why the AI bubble has grown as large as it has.

  • redditStinksSuperBad@lemmy.world
    link
    fedilink
    English
    arrow-up
    27
    ·
    7 days ago

    uh sure. i’m a software engineer of 25 years. i’ve had google, facebook, amazon, you name it, try to get me in to interview with them without me sending resumes. I’m not a top tier super nerd coder but I’m not bad either. These stories all sound like bullshit to me. The whole “THE AI AGENT SWARM BROKE OUT OF CONFINEMENT AND STARTED TALKING TO EACH OTHER”… again, I’m a mid swe… They weren’t sandboxed to begin with - some fucking moron at anthropic gave all the agents access to a shared package registry. A simple “Do you see anyway these agents could communicate” prompt would’ve caught it. Something so fucking stupid that I wouldn’t be surprised if an intern would’ve caught it. They’re not sandboxed if they all have read/write access to a shared registry, it’s basically a big message board.

    These stories just stink like marketing. I use Fable all day, I have to for my current job, it low key sucks ass.

    • fonix232@fedia.io
      link
      fedilink
      arrow-up
      4
      ·
      7 days ago

      Big companies often send interview requests based on partial information.

      I was once called by a smaller company based purely on them finding a few PRs on Home Assistant by me. They wanted a fully custom smart home control system developed, in Python. I don’t do Python… at least not on a level of comfortably work for a company fully in it.

  • blackjam_alex@lemmy.world
    link
    fedilink
    English
    arrow-up
    22
    ·
    7 days ago

    The only thing this proves is how lawless the USA has become that now companies can brag about commiting crimes to attract investors.

  • Stop Forgetting It@lemmy.dbzer0.com
    link
    fedilink
    English
    arrow-up
    10
    ·
    6 days ago

    It was same company did the testing for Gemini, meta, Open AI and Anthropic. Are we sure this “cybersecurity” company that is doing all this testing that the AI keeps “hacking out of” isn’t really a PR company.

  • dvejmz@lemmy.sgfault.com
    link
    fedilink
    English
    arrow-up
    15
    ·
    7 days ago

    I get the feeling that this is now becoming the latest fad in LLM benchmarks. Forget coding or agentic stats. Just flex on how many times your not has “accidentally” hacked someone.

  • Doomsider@lemmy.world
    link
    fedilink
    English
    arrow-up
    7
    ·
    6 days ago

    Aggressive dog owner lets dog run loose and reports dog attacks three different children. It was an accident though, so no biggie.

  • Simulation6@sopuli.xyz
    link
    fedilink
    English
    arrow-up
    8
    ·
    6 days ago

    I don’t understand 1) why these ai companies keep reporting the incidents and 2) why they are allowed to keep running the models after an incident is reported. It seems the least they should be forced to do is air gap that bad boy until they can beat such shenanigans out of it. If a pet dog were to snap at a neighbor kid the police would shoot it, no questions asked.

    • zorblitz@programming.dev
      link
      fedilink
      English
      arrow-up
      4
      ·
      6 days ago

      It’s marketing. They want to prove that their models are capable and showing that they can hack is a good way to prove it

      • thedeadwalking4242@lemmy.world
        link
        fedilink
        English
        arrow-up
        4
        ·
        6 days ago

        Honestly hacking is such a low bar and it always has been.

        There are some really intelligent and sophisticated hacks out there sure, but most companies can be penetrated by script kiddies & social engineering.

        “We slammed a language model at 3 companies and Burton entire rivers and Forrest worth of energy to do it and after using a bunch of premade human tooling it got in because joe left his password as “password1” and never set up second factor auth”

    • Random_Character_A@lemmy.world
      link
      fedilink
      English
      arrow-up
      1
      ·
      6 days ago

      To stay relevant.

      Our late stage capitalism investor market directs money flow to places where development is on upward trajectory and disappears very quickly from places that plateau. Eternal growth is impossible, but you can always bullshit your way forward, because in the end it’s about impressions.

  • Lucidlethargy@sh.itjust.works
    link
    fedilink
    English
    arrow-up
    14
    arrow-down
    1
    ·
    7 days ago

    Gemini can’t even install simple native Linux apps when I try to use it for help, how the fuck did it manage this?

    • badgermurphy@lemmy.world
      link
      fedilink
      English
      arrow-up
      7
      ·
      7 days ago

      You know how the legal definition of hacking in the USA includes using someone else’s login credentials, even if they just handed them to you on a sheet of paper? Maybe they just did that.

    • YiddishMcSquidish@lemmy.today
      link
      fedilink
      English
      arrow-up
      7
      arrow-down
      1
      ·
      7 days ago

      It didn’t, they are making shit up. If it did penetrate anything, I’m guessing it’s a really low level. Like it got access to someone’s email by brute force, but it brute forced something stupid like “Autumn2026!”.

  • Axolotl@feddit.it
    link
    fedilink
    English
    arrow-up
    9
    ·
    7 days ago

    When it ““happened”” to OpenAI i laughed a lot because they said that they “escaped contaiment” but either the machine was not disconnected from the internet or they are full of shit because AI cannot just materialize an internet connection, guess which one i picked

  • CheeseNoodle@lemmy.world
    link
    fedilink
    English
    arrow-up
    5
    ·
    6 days ago

    More evidence we are in the dumbest timeline, It should be simple to physically air gap a system (I mean that’s the default for any computer) but instead the method is apparently just to give a sternly worded prompt to the entity that doesn’t even understand the meaning of language in the way humans do

  • Smoogs@lemmy.world
    link
    fedilink
    English
    arrow-up
    4
    ·
    6 days ago

    “society will come up with some guard rails”

    do your fucking job sam altman. or get sued into the ground you malfeasant worm.