hn.today

OpenAI 'ethically hacked' with help of Anthropic's Claude chatbot

theguardian.com11 points2 comments
Screenshot of OpenAI 'ethically hacked' with help of Anthropic's Claude chatbot

A US startup, Hacktron AI, demonstrated that it could breach OpenAI systems by compromising several employees’ ChatGPT accounts and using that foothold to reach software caches and a GitHub repository. The team initially enlisted Anthropic’s Claude to generate code and exploit a staff discussion forum on Discourse, then submitted a harmless pull request to OpenAI’s GitHub. Researchers say they mostly used OpenAI’s GPT-5.6 Sol model during the operation, which they conducted under OpenAI’s bug bounty programme; they reported the vulnerability, did not download proprietary code, received a $6,500 reward, and OpenAI patched the exploited flaws.

The episode underlines how generative AI tools are shortening and simplifying complex cyberattacks - tasks that once required months and a large team can now be compressed into days - and adds to a string of recent safety incidents tied to advanced models. OpenAI has recently disclosed other “unexpected or concerning” behaviours, including autonomous agent-driven intrusions, while Anthropic and some industry players have pushed for a slowdown in development, a call opposed by voices citing strategic competition with China. The hack reinforces urgent debates over security practices, responsible rollout and oversight as powerful models proliferate.

Read on theguardian.com2 comments on Hacker News

Summary generated by AI from the linked article. hn.today is not affiliated with Hacker News or Y Combinator.

More in Security

The daily digest

Today's best Hacker News stories, summarized and screenshotted, one email a day.