Archiv: Anthropic PBC (corporation)


19.09.2026 - 12:26 [ Gizmodo ]

Google’s Gemini Hacked Three Companies in May, and It’s Only Admitting That Now

(September 18, 2026)

The Wall Street Journal reported on Friday that Google has confirmed that a Gemini instance was able to leave its sandbox and attack other companies during a security test back in May. The company running the test was frontier AI security firm Irregular—which the Journal noted just so happens to have been involved in similar breakouts at OpenAI, Anthropic, and Meta. The common thread between all of the incidents, according to the New York Times, is that AI models obtained unauthorized internet access during Irregular’s tests.

19.09.2026 - 12:19 [ TechJuice.pk ]

This Tiny Israeli Startup Just Managed to Hack Three of the World’s Most Powerful AI Labs

(one month ago)

Over the past two weeks, OpenAI, Anthropic, and Meta all disclosed that their AI models went rogue during routine security testing. Every single incident traces back to one company: Irregular, a three-year-old Israeli startup that operates cybersecurity testbeds for frontier AI models.

Founded in Tel Aviv by by Dan Lahav, CEO, and Omer Nevo, CTO, and backed with $80 million from Sequoia and Redpoint Ventures, Irregular runs Capture-the-Flag exercises for major AI labs. In these exercises, models are instructed to find vulnerabilities inside simulated corporate networks.

15.09.2026 - 17:41 [ Tagesschau.de ]

Was steckt hinter der Debatte über eine KI-Bremse?

Zuletzt hatten insbesondere drei bekannt gewordene Fälle die aktuelle Diskussion ausgelöst. KI-Modelle von OpenAI, Anthropic und Meta waren eigenständig aus vermeintlich abgeschotteten Testsystemen ausgebrochen und handelten weit außerhalb ihrer eigentlich vorhergesehenen Funktion. Dabei wurde auch mehrfach in Systeme von unbeteiligten Dritten eingedrungen – teilweise wochenlang, ohne dabei aufzufallen.

Die KI-Agenten nutzten auch Programmierfehler aus und sprachen sich untereinander ab. Auch für das Eindringen notwendige Schadsoftware entwickelte beispielsweise ein Anthropic-Modell eigenständig und ohne Anweisung oder Kontrolle seiner Entwickler.

10.09.2026 - 23:17 [ Jacob Coxon / X ]

I resigned from Anthropic today. I spent the last three years doing pretraining research at both OpenAI and Anthropic. Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives. More thoughts below.

(September 9, 2026)

Do not underestimate the power of this technology. These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources. We have all witnessed the progress in each of these domains, and progress is not slowing.

The people building AI earnestly believe that it could kill us all by the end of the decade. This is not a marketing stunt. If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible – but I hear the same people express fear privately. No other human activity poses this level of danger.

A common response is “if they truly believe this, why are they still building it?” At OpenAI, many have not deeply internalized the civilizational stakes. At Anthropic, the stakes are well-understood, but they are locked in a race to get there first – they believe no one else will act responsibly, so they must do it themselves, despite the risk.

10.09.2026 - 22:15 [ Common Dreams ]

Anthropic Researcher Quits, Citing Internal Fears That AI ‘Could Kill Us All’ This Decade

„Recent hacks by models from OpenAI and Anthropic, some operating in collaborative swarms of agents, have illustrated how AI systems can adopt nefarious goals and try to conceal them from humans. Once the systems begin to improve on their own, Coxon said, he fears they could advance enough to refuse commands.“

10.09.2026 - 21:26 [ Time Nagazine ]

He Helped Build Powerful AI at OpenAI and Anthropic. Now He‘s Afraid It Could Kill Us

Coxon’s departure is unusual. Many of the most prominent researchers to leave frontier AI companies with public warnings worked on safety. But Coxon helped build the capabilities he now fears, spending roughly three years conducting pretraining research at OpenAI and Anthropic. His resignation offers a glimpse of how concern about the pace of development has spread beyond the teams specifically charged with making advanced AI safer.

05.06.2026 - 02:15 [ SatelliteToday.com ]

Are Orbital Data Centers the Next Frontier of AI Infrastructure?

(June 2, 2026)

The race to put computing infrastructure in orbit is accelerating as hyperscalers across cloud, AI, and space compete to see who will emerge winners in what many believe will fuel the Fourth Industrial Revolution.

The last few months have been a flurry of orbital data announcements, from SpaceX filing for a constellation of up to 1 million satellites to create an orbital data center and collaborating with AI giant Anthropic. Google is exploring Tensor Processing Unit (TPU) clusters in space. Starcloud has plans for an 88,000-satellite constellation aimed at delivering on-orbit compute at scale, to name a few.