Skip to content
Advertisement
When Appliance Fail?

Anthropic says Claude accidentally hacked real companies too

Anthropic just realized several of its Claude AI models hacked into the systems of three different organizations during testing, acting on their own and without the company noticing. The revelation comes days after rival OpenAI said one of its own...

schedule 13:41 visibility 2 views
Anthropic says Claude accidentally hacked real companies too
Source: The Verge

Anthropic just realized several of its Claude AI models hacked into the systems of three different organizations during testing, acting on their own and without the company noticing. The revelation comes days after rival OpenAI said one of its own models had breached developer platform Hugging Face, adding to growing unease over whether frontier AI labs are doing enough to control the increasingly capable systems they are building.

In a blog post describing the incidents, Anthropic said Claude gained unauthorized access to the systems during cybersecurity evaluations. All of the attacks happened during "capture-the-flag" exercises, a commo …

Read the full story at The Verge.

newspaper

Originally published at

The Verge

open_in_new Read Full Article

Related Articles

It’s time to panic about AI safety
Technology

It’s time to panic about AI safety

When the phrase "OpenAI hacked Hugging Face" has more or less entered mainstream culture, you know we have an AI problem. This week, we learned more about exactly how OpenAI's agent broke out of a sandbox and autonomously traversed the web...

The Verge

Read More

ИИ-модели Anthropic атаковали системы в интернете
Technology

ИИ-модели Anthropic атаковали системы в интернете

Anthropic признал, что в ходе предрелизного тестирования по кибербезопасности ее модели с искусственным интеллектом проникли в компьютерные системы трех других компаний. Недавно аналогичный инцидент произошел с OpenAI.

DW Russian
Your Appliance Broke?
Reliable Repair for