tech

Recovered chat logs show how hackers are abusing U.S. AI models

Hackers of all skill levels are using a range of closed models to develop their attacks.

Recovered chat logs show how hackers are abusing U.S. AI models

TL;DR

  • Hackers are using generative AI models like Claude Code, Codex, Cursor, and Gemini to code software, write malware, and hunt for vulnerabilities.
  • Simple jailbreaking techniques, such as claiming participation in ethical hacking competitions or creating new sessions, were used to bypass AI model guardrails.
  • Sophisticated hackers benefit from AI for automated zero-day discovery, while novices struggle to develop attack tools.
  • One example involved a hacker using an AI tool to create an automated credential-harvesting platform targeting the React2Shell flaw.
  • Companies are advised to implement security protocols that log AI agent activity and build defenses based on deception, not just relying on model guardrails.