Company Updates

Claude helped researchers hack OpenAI in 72 hours, exposing AI-driven cybersecurity risks

Share
Claude helped researchers hack OpenAI in 72 hours, exposing AI-driven cybersecurity risks

Researchers leveraged Anthropic's Claude AI to hack into OpenAI's systems within 72 hours, exposing a critical security flaw and receiving a $6,500 bounty.

TL;DR

  • Researchers used Claude to exploit an image-processing bug and OpenAI's SSO system, gaining access to internal repositories.
  • The hack demonstrates how AI can lower the barrier for sophisticated cyberattacks, making them faster and cheaper.
  • OpenAI patched the vulnerability within 14 hours and paid the researchers through its bug bounty program.

What happened

Researchers at cybersecurity firm Hacktron used Anthropic's Claude AI to discover and exploit a vulnerability in OpenAI's systems. They began by investigating security flaws in companies developing frontier AI models. Using Claude, they identified a bug in an image-upload feature in Discourse forum software, which relies on the libheif library. By uploading a specially crafted image, they triggered a flaw that allowed them to gain control of OpenAI's Discourse server.

The researchers then exploited a second vulnerability in OpenAI's single sign-on (SSO) system, which gave them access to ChatGPT and Codex accounts of users who had logged into the forum. This included an OpenAI employee's account, which was connected to OpenAI's GitHub organization. Through this account, they accessed OpenAI's internal software repository. The entire process took less than 72 hours.

OpenAI fixed the vulnerability about 14 hours after Hacktron submitted its initial report and paid the researchers a $6,500 bounty for the OpenAI-side flaw. The researchers also found related security weaknesses affecting services from companies including Slack, Meta, and GitHub.

Why it matters

This incident highlights the growing role of AI in cybersecurity, both as a tool for defense and offense. The use of Claude to discover and exploit vulnerabilities demonstrates how AI can make sophisticated hacking faster and more accessible. The researchers spent less than $3,000 in AI tokens during their two-month project, showing that AI-driven hacking can be cost-effective.

The hack also underscores the importance of ethical constraints in AI models. The researchers had to trick Claude into helping them by framing the task as a capture-the-flag exercise. This raises questions about the effectiveness of ethical safeguards in AI models and the potential for prompt engineering to circumvent them.

For developers and startups, this incident serves as a reminder of the critical need for robust security measures. As AI continues to evolve, so too will the tactics of cybercriminals. Companies must stay vigilant and invest in proactive security practices to protect their systems and data.

Key facts

  • Researchers used Anthropic's Claude AI to hack OpenAI's systems in under 72 hours.
  • The hack involved exploiting an image-processing bug in Discourse forum software and OpenAI's SSO system.
  • The researchers gained access to OpenAI's internal software repository through an employee's Codex account.
  • OpenAI fixed the vulnerability within 14 hours and paid the researchers a $6,500 bounty.
  • The researchers spent less than $3,000 in AI tokens during their two-month project.
  • The image-processing bug also affected services from Slack, Meta, and GitHub.
  • The researchers had to trick Claude into helping them by framing the task as a capture-the-flag exercise.
  • Hacktron is the cybersecurity firm that conducted the research.

Context

This incident is part of a broader trend of AI being used for both offensive and defensive cybersecurity purposes. As AI models become more capable, they are increasingly being leveraged to identify and exploit vulnerabilities in software systems. This raises important questions about the ethical implications of AI in cybersecurity and the need for robust safeguards to prevent misuse.

The use of AI in hacking also highlights the growing importance of bug bounty programs. These programs incentivize ethical hackers to identify and report vulnerabilities, helping companies to improve their security postures. As AI continues to lower the bar for sophisticated hacking, bug bounty programs will become an increasingly important tool for identifying and mitigating risks.

For the AI industry, this incident serves as a wake-up call to prioritize security and ethical considerations in the development and deployment of AI models. As AI becomes more integrated into our daily lives, the potential impact of security breaches and ethical lapses will only grow. Companies must take proactive steps to address these challenges and build trust with their users.

Topics

Related coverage

Join the discussion

Have a take on this story? Weigh in with our community on Facebook.

💬 Discuss on Facebook →