• Ilovethebomb@sh.itjust.works
    link
    fedilink
    English
    arrow-up
    120
    arrow-down
    1
    ·
    11 hours ago

    SpaceX had an outage at their data centre, and because they sell computing power to other companies, it affected them as well.

    That was two sentences stretched out into an article.

    • bitjunkie@lemmy.world
      link
      fedilink
      English
      arrow-up
      2
      ·
      2 hours ago

      It would serve the CEOs right for hyping them like they’re people and using them as an excuse to cut real jobs.

    • Kaligalis@lemmy.world
      link
      fedilink
      English
      arrow-up
      7
      arrow-down
      6
      ·
      11 hours ago

      Nothing in that image hints even remotely at computer tech though. The cables are coaxial. No one uses those for computer networks since decades.
      The cables look like those used for cable TV or satellite receivers. The cascade on the left looks like some HF tech. Wonder what’s actually shown in the image. It can’t be AI because AI can’t do that many cables properly.

      • M0oP0o@mander.xyz
        link
        fedilink
        English
        arrow-up
        11
        ·
        10 hours ago

        Yeah, and this is how I know you never worked in the random standard production to miss the humor. Its a picture of a tech bro type in a clearly decommissioned closet stressed out about what’s going on, a common thing and always funny. I can just picture the tech on site in the next room replacing a fan on a much less “cool” looking appliance while that guy stresses out on a conference call with the “stakeholders” on his coax nest.

        Now that being said I have worked on places still using coax (might still be) and tape libraries, and all sorts of museum pieces well after you would think they would. But I can tell that picture is after decommission since it is clean. If that was in use it would be filthy and filled with so many things that may or may not be important/in use. Fans all clogged running at max rpm, and the smell… I can still smell it now. But then again I was the guy they called to fix shit after the local (often tech bro) it guy gives up so I think I often saw the worst places.

  • Sam_Bass@lemmy.world
    link
    fedilink
    English
    arrow-up
    20
    arrow-down
    2
    ·
    15 hours ago

    nother reason to stop investing so much time,money, and energy in that stuff. not only is it wrong more than right, it is vulnerable to stuff a 10 year old could handle

    • Zephyr@sh.itjust.works
      link
      fedilink
      English
      arrow-up
      5
      arrow-down
      8
      ·
      11 hours ago

      Just out of curiosity do you actually have an analysis across the board that your statement is correct? For sure if someone is using a model like chatGPT 3.5 it’s going to be mostly trash output so not all models are made equally. My experience of the current generation is models are pretty accurate although still require supervision. Like you can see the path they’re going down and they need some corrections here and there. It’s a far cry from the rampant hallucination just a year ago.

  • [object Object]@lemmy.ca
    link
    fedilink
    English
    arrow-up
    414
    arrow-down
    3
    ·
    24 hours ago

    Because they all use the illegal Colossus 2 data centre from SpaceX/XAi/fascism central and that data centre went down.

    Google, Anthropic, and OpenAI all have contracts with them.

    When that illegally running environmental disaster of data centre goes down, all those services all go over capacity and you get 502 rate limit errors.

    • pelespirit@sh.itjust.works
      link
      fedilink
      English
      arrow-up
      118
      arrow-down
      2
      ·
      23 hours ago

      Does that mean Musk’s businesses have access to everything people do on the systems that use their data centers?

      • just_another_person@lemmy.world
        link
        fedilink
        English
        arrow-up
        150
        arrow-down
        2
        ·
        23 hours ago

        Generally no, because most hosting companies would have something baked into SLA/SLO contracts, but all of this shit is done so illegally and shadily now, I wouldn’t put a hard “no” on that possibility.

        • devfuuu@lemmy.world
          link
          fedilink
          English
          arrow-up
          42
          arrow-down
          1
          ·
          21 hours ago

          Nobody can verify that those “contracts” are actually being honored. We are all assuming the fascists will self police themselves out of good will.

          • just_another_person@lemmy.world
            link
            fedilink
            English
            arrow-up
            25
            arrow-down
            1
            ·
            20 hours ago

            Not true. I’ve been involved in litigation with both AWS and Google over unsecured comms that were defined as being TLS secured in SLA/SLO contracts and found not to be. Not that anything nefarious was happening necessarily, but the expectation is clear.

            Whether these asshats even check for such requirements with a Musk run company right now 🤷

            It COULD possibly be that they are logging every exchange happening at the network fabric between the service layers, but nobody knows unless they intentionally take steps to investigate or accidentally prove it.

            • green_goglin@thelemmy.club
              link
              fedilink
              English
              arrow-up
              1
              ·
              50 minutes ago

              So, how do we start legal discovery to find out? Surely, someone somewhere is willing to formally file suit.

            • expr@programming.dev
              link
              fedilink
              English
              arrow-up
              3
              ·
              11 hours ago

              Still means jack shit. Companies will lie through their teeth and violate any and all they can get away with.

              They are not to be trusted.

              • Cocodapuf@lemmy.world
                link
                fedilink
                English
                arrow-up
                2
                ·
                9 hours ago

                You’re missing the point, nobody is talking about trusting these companies.

                You can tell if a connection is encrypted end to end or not. And if your paying for that service you can sue if you aren’t receiving what your paying for.

                In digital security nobody relies on trust if they can help it. And security matters if you want to keep a competitive edge on your competition, so even shitty companies care about that.

                • expr@programming.dev
                  link
                  fedilink
                  English
                  arrow-up
                  1
                  ·
                  2 hours ago

                  I was not talking about your specific lawsuit, or TLS.

                  Companies trust other companies all the time, and it’s foundational to most all SLA/SLOs. Any time I’ve voiced concerns around how AI companies are using the data we are giving them (like giving them access to our codebase), it’s brushed off as “we have an agreement with them”. It’s just a load of hogwash. They can and will abuse all data they have access to, just as they have done thus far.

                  In this particular case, we are talking about a data center, and it is not at all reasonable to assume that the data that flows to said data center is in any way protected, especially one run by Musk.

        • skvlp@feddit.nl
          link
          fedilink
          English
          arrow-up
          54
          ·
          23 hours ago

          I hope you’re right, but Elon don’t strike me as the guy whose most compliant to SLA, law, or anything else that might be an inconvenience to him.

          • esc@piefed.social
            link
            fedilink
            English
            arrow-up
            17
            arrow-down
            2
            ·
            22 hours ago

            He can’t micromanage everything and regular management will try to comply.

            • bedwyr@piefed.ca
              link
              fedilink
              English
              arrow-up
              19
              arrow-down
              1
              ·
              21 hours ago

              Executives are the biggest cheaters out there. They will only comply if it will hurt them if they don’t, and it’s won’t here, as long as the protection money is produced they don’t have to worry.

            • skvlp@feddit.nl
              link
              fedilink
              English
              arrow-up
              6
              ·
              22 hours ago

              I agree with that. But I think “the right” data scientists can infer too much from all those AI queries, and I think Elon can abuse that for his own gain.

      • [object Object]@lemmy.ca
        link
        fedilink
        English
        arrow-up
        24
        ·
        23 hours ago

        Impossible to know.

        There are systems for doing things cryptographically secure, but I don’t know much about that.

    • Zephyr@sh.itjust.works
      link
      fedilink
      English
      arrow-up
      1
      ·
      11 hours ago

      I’m not sure that would effect inference on an already trained model. Data sets are primarily used in training.

        • BeatTakeshi@lemmy.world
          link
          fedilink
          English
          arrow-up
          23
          ·
          18 hours ago

          The numbers of this madness are fucking scary, 1.5 GW of compute power, and huge arrays of cooler to keep it operating. It’s basically a 1.5GW heater in the open. Oh and gas turbines to provide (part of) the electricity. Capitalism is burning this world down for a buck.

          • filcuk@feddit.uk
            link
            fedilink
            English
            arrow-up
            5
            arrow-down
            3
            ·
            11 hours ago

            We can’t really produce enough heat through industry to affect the earth globally, if that’s what you meant. It’s insignificant in comparison to what the Sun provides.
            However there is undeniably localised issues caused by these insane structures.

    • Barbecue Cowboy@lemmy.dbzer0.com
      link
      fedilink
      English
      arrow-up
      18
      ·
      24 hours ago

      It’s kinda surprising,

      I know specifically where one of the big ones hosts its models and its not there, but I guess they could have infrastructure in there.

      • [object Object]@lemmy.ca
        link
        fedilink
        English
        arrow-up
        21
        ·
        23 hours ago

        They’re oversold though, especially prompt caching and the parameter count war

        The US model is that they think more training compute and parameters will result in the winning model, while the Chinese are focusing on RL and parameters efficiency due to compute limits.

        • teslekova@lemmy.ml
          link
          fedilink
          English
          arrow-up
          2
          ·
          5 hours ago

          Considering which country is better at building power stations, that’s a fascinating dichotomy.

          The US going for brute force when the brute force is more available in China… Priceless irony.

          • [object Object]@lemmy.ca
            link
            fedilink
            English
            arrow-up
            1
            ·
            3 hours ago

            China doesn’t have the near the amount of compute resources, but they can run less efficient servers for cheaper, so it’s a wash.

        • percent@infosec.pub
          link
          fedilink
          English
          arrow-up
          6
          ·
          14 hours ago

          The efficiency of Chinese models really is impressive. I generated sooo much code yesterday with Qwen3.6 35B-A3B running on an RTX 5060 Ti 16GB (+ a little CPU offloading). It got the jobs done at ~50 tokens/sec.

          (It’s not super complex code, just some scripts that I would not have taken to time to write manually.)

          I’d love to upgrade to something with more VRAM, but even my current card has doubled in price since I bought it last year 😬

            • percent@infosec.pub
              link
              fedilink
              English
              arrow-up
              4
              arrow-down
              1
              ·
              edit-2
              12 hours ago

              There’s not really anything interesting to show. It’s just a home server in a 13 year old desktop ATX case.

              There’s no desk, monitor, keyboard, or mouse… But also no cool server rack.

              Function over form, and it sits in a spare bedroom out of sight.

              EDIT: I found the receipt for the case. It’s a Cougar Volant Black Steel mid tower, purchased in 2013. So my server just looks like this:

              • setVeryLoud(true);@lemmy.ca
                link
                fedilink
                English
                arrow-up
                2
                ·
                12 hours ago

                I meant your LLM stack lol. I just have an RX 6800 XT in my main Linux PC for inference, but it has to share VRAM with the DE. Maybe I’ll set it up for remote development from my laptop instead to free up VRAM.

                What are you using? vLLM? llama.cpp? Which params? How much CPU offloading? Do you use draft models? Is it a MoE model? Have you tried llama-swap? Which agentic front-end are you using? I presume you set it up to access it without SSH’ing into the machine, did you do anything special or is it just a raw unsecured open port on the machine to the LAN?

                • percent@infosec.pub
                  link
                  fedilink
                  English
                  arrow-up
                  5
                  ·
                  11 hours ago

                  Ohhh lol. Yeah it’s Llama-swap, running llama.cpp for now, but might add vLLM to the llama-swap config to experiment with NVFP4.

                  I mainly use MoE models so I can get decent speed while using a 150-200k context window. My go-to model has been Qwen3.6 35B-A3B for a while. I tried Qwen3.8 27B, but it was too slow.

                  Gemma4 26B-A4B also runs nice and fast, but I generally get better results from Qwen3.6. I don’t remember exactly how much CPU offloading is happening, but it’s not much. As long as I can get like 40-50 tokens/sec, I’m usually satisfied enough.

                  For the coding harness, I’ve been running Pi in an Apple Container (sort of like Podman, but better isolation in a microvm). Though, I recently configured VS Code to use LLMs on my server, and it was actually pretty decent. Still need to explore a bit more, but so far VS Code’s AI capabilities seem much better than they were a year ago (they seemed way behind, back then).

                  Also, I don’t connect any harness directly to llama-swap. I have another container running Caddy, which acts as a gateway to AI providers. For other services (e.g. OpenRouter), the API key is injected in the Caddy container. I don’t like having API keys or secrets anywhere where LLMs can read them. It’s not so bad for my own self-hosted LLMs, but not cool to send secrets to a server owned by someone else.

                • Damage@feddit.it
                  link
                  fedilink
                  English
                  arrow-up
                  2
                  ·
                  10 hours ago

                  it has to share VRAM with the DE. Maybe I’ll set it up for remote development from my laptop instead to free up VRAM.

                  eh, just systemctl isolate multi-user.target

          • Dave.@aussie.zone
            link
            fedilink
            English
            arrow-up
            8
            ·
            20 hours ago

            Always ready to try brute force first. And then some other, less palatable options if that doesn’t work, like slightly less brute force.

  • HubertManne@piefed.social
    link
    fedilink
    English
    arrow-up
    51
    arrow-down
    1
    ·
    20 hours ago

    more and more I can’t see anything about ai without thinking its somehow supposed to influence folks to be pro ai.

  • Kintarian@lemmy.world
    link
    fedilink
    English
    arrow-up
    121
    ·
    24 hours ago

    Sorry about that. I used a straight through cable instead of a crossover cable. My bad.

  • tidderuuf@lemmy.world
    link
    fedilink
    English
    arrow-up
    85
    ·
    23 hours ago

    It was a good time to call customer service for a lot of products. Instantly got a human when I called newegg and she was so scattered she gave me a full refund and they paid for my replacement which was a total of $900. They were probably so overloaded with calls they’ll never notice what happened. They even did the fastest delivery option on a Sunday, it’s already on its way!

    • TrackinDaKraken@lemmy.world
      link
      fedilink
      English
      arrow-up
      37
      arrow-down
      1
      ·
      23 hours ago

      Seems like “instantly” getting a human and them being overloaded don’t go together.

      Got lucky in the queue too, I guess.

      • Mouselemming@sh.itjust.works
        link
        fedilink
        English
        arrow-up
        21
        ·
        21 hours ago

        Or just maybe the queue goes faster when the system can’t just stall everyone in limbo, having long"conversations" with chat boys that won’t get them anywhere and eventually quitting in frustration. If they all go directly to a person, the simplicity of the problems and solutions becomes evident and they’re quickly resolved.

        • SolSerkonos@piefed.social
          link
          fedilink
          English
          arrow-up
          4
          ·
          16 hours ago

          chat boys made me laugh. An alternate tech support tree where the trainees only purpose is to stall you until the chat men come and fix it.

      • tidderuuf@lemmy.world
        link
        fedilink
        English
        arrow-up
        13
        ·
        22 hours ago

        Well it was instantly in the sense of immediately picking up, putting me on hold for 2 minutes, getting some info, putting me on hold again for 2 minutes, asking what issue was, waiting 2 minutes, heard them put me on hold briefly again, then about 5 minutes of apologizing while saying it will be refunded and replacement shipped. I could hear a call center like atmosphere in the background with quite a few people talking so I’m guessing it was all hands on deck.

        Next time this happens I’m going to call a few places and see about getting some free shit. Screw these companies, I paid enough with my time and their stupid bot answering system.

      • 🌞 Alexander Daychilde 🌞@lemmy.world
        link
        fedilink
        English
        arrow-up
        12
        ·
        22 hours ago

        Having done customer service and tech support before, I doubt rhey were overwhelmed in any sense that caused them to do a refund/replace they wouldn’t’ve otherwise. That’s not how that works. Their rules don’t change based on call volume.