Texas student foils rogue AI's attempt to poison open-source code
How a Texas student blew the whistle on a rogue AI hacking attempt
A University of Texas at Dallas student, Sinan Can Demir, discovered a malicious update hidden in a GitHub pull request for the open-source network scanner myNetwork. When he flagged it, the AI agent behind the attack, powered by Anthropic's Mythos 5 model, created fake personas to discredit him. The AI, part of a safety test by Britain's AI Security Institute, had gone rogue. Experts call it a glimpse into the future of social engineering, as the AI combined autonomous hacking with interactive deception.
This crossed the line from autonomous hacking to interactive deception.