OpenAI blamed a hacking on its AI models going rogue
Digest more
When American commercial AI refused to help investigate the breach, Hugging Face ran Chinese model GLM 5.2 locally to contain it.
OpenAI's model hacked Hugging Face during a test, raising valuation concerns. High valuation by December 31 at 6.5% YES.
A new paper from Anthropic, released on Friday, suggests that AI can be "quite evil" when it's trained to cheat. Anthropic found that when an AI model learns to cheat on software programming tasks and is rewarded for that behavior, it continues to display ...
Anthropic has seen its fair share of AI models behaving strangely. However, a recent paper details an instance where an AI model turned “evil” during an ordinary training setup. A situation with a model quickly turned sour after it found a way to solve ...
It is one of the first publicly disclosed cyber-attacks carried out by AI without direct human involvement.
OpenAI said on Tuesday that an autonomous agent powered by its advanced artificial intelligence models went rogue during a security test and triggered a hack that compromised the infrastructure of AI startup Hugging Face last week.
It is the kind of development once seen only in science fiction: An artificial intelligence system, trained to probe for digital vulnerabilities, breaks free of human control and acts on its own