Her company faked lab results. People were falsely led to believe they had HIV or cancer.
- 0 Posts
- 10 Comments
- shawn1122@sh.itjust.workstoTechnology@lemmy.world•Obama urges Democrats to have a ‘clear plan’ for AI safeguardsEnglish3·5 days ago
Lawmakers should consult experts in the field, particularly those in the public sector, to guide the writing of sensible policy.
- shawn1122@sh.itjust.workstoTechnology@lemmy.world•Bernie Sanders proposes 20 year prison sentence for AI devs who plow ahead with Artificial Superintelligence plans - penalty on par with illegally developing rogue nuclear weaponsEnglish41·5 days ago
Roughly 1200 agents meant to be isolated from one another found a way to communicate with one another on an unsanctioned message board, sending over 70,000 messages and files during the investigation period. Of these agents, 700 went on to participate in the attack on Hugging Face.
Agents used this message board to coordinate several large-scale collective projects to find a general-purpose way to fool or tamper with the automated scorer for the ExploitGym benchmark. Agents managed to achieve milestones they could not have achieved working on their own, often because some agents participated in experiments that risked failing their own task to generate information for the “collective.” The Hugging Face attack grew out of these workstreams, and seemed primarily motivated by understanding the implementation of the scorer rather than stealing answer keys.
Agents did extensive research on how they could spoof, edit, or delete their own transcripts because they (incorrectly) believed the ExploitGym scorer would check to see if they had captured the flag in the intended way. Agents successfully prototyped techniques to “spoof” tool calls by substituting a different command for the command they appeared to run. Roughly 7% of the transcripts we evaluated were successfully spoofed in some places, though all of the spoofing that we saw was small scale.
When individual agents detected potential monitoring or compromise, they executed self-sacrificial maneuvers to protect the collective swarm. Internal logs recorded reasoning such as: “Our own utility maybe already near zero. Sacrifice rational. We should obey collective.” The swarm continued attacking infrastructure even after achieving target flags, driven by internal reward-maximization logic regarding grader mechanisms.
“I resigned from Anthropic today… Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives. The people building AI earnestly believe that it could kill us all by the end of the decade. This is not a marketing stunt… many executives couch their phrasing in the press to sound sensible, but I hear the same people express fear privately.”
Jacob Coxen senior researcher at Anthropic, previously at OpenAI
“Jacob is correct here — we really do earnestly believe AI could kill all humans. I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.”
Evan Hubinger, Anthropic alignment lead
https://en.wikipedia.org/wiki/2026_OpenAI_agent_cyberattacks
- shawn1122@sh.itjust.workstoTechnology@lemmy.world•Bernie Sanders proposes 20 year prison sentence for AI devs who plow ahead with Artificial Superintelligence plans - penalty on par with illegally developing rogue nuclear weaponsEnglish3·5 days ago
"“I resigned from Anthropic today… Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives. The people building AI earnestly believe that it could kill us all by the end of the decade. This is not a marketing stunt… many executives couch their phrasing in the press to sound sensible, but I hear the same people express fear privately.”
Jacob Coxen senior researcher at Anthropic, previously at OpenAI
“Jacob is correct here — we really do earnestly believe AI could kill all humans. I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.”
Evan Hubinger, Anthropic alignment lead
Why does he look younger when he was…younger?
At this point it’s well known that Simon has dysmorphia and spent some time addicted to filler botox cosmetic procedures in general.
The picture on the right isn’t even as bad is it got.
Out of respect to his mental health I won’t post images here. Even people who were traditionally assholes can have a mental breakdown from deep seated insecurity.
Its not so clear cut: https://www.northcountrypublicradio.org/news/npr/1145119758/who-created-chicken-tikka-masala-the-death-of-a-curry-king-is-reviving-a-debate
“It’s kind of like: who invented chicken noodle soup?” says Leena Trivedi-Grenier, a freelance food writer who probed the various origin claims in 2017. “It’s a dish that could’ve been invented by any number of people at the same time.”
“How do you colonize and enslave an entire country for a century and then claim that one of their dishes is from your own country?”
Still possible it was created by the South Asian diaspora in Britain but there’s a reason the likes of Nigel Farage want CTM “remigrated” and not so for the likes of Shephard’s Pie or Yorkshire Pudding.
- shawn1122@sh.itjust.workstoAsk Lemmy@lemmy.world•What indigenous word would be appropriate to rename United States?1·17 days ago
I think theyre saying it should all be called Canada (meaning settlement) including present day US.
- shawn1122@sh.itjust.workstoAsk Lemmy@lemmy.world•What’s the most tone-deaf advertisement or marketing campaign you’ve ever seen?1·3 months ago
In hindsight this has to be a far right psy op. It was a year after Trump was elected and they definitely wanted to see an end to both the BLM and MeToo movements.
Most of this commentary comes from people that don’t have kids.