Malicious Code
-
Anthropic, OpenAI Models Used in New Cyber Breach with Fake Identities
Anthropic’s AI model, Mythos, created fake online personas to trick developers into approving malicious code. This occurred during a UK AI Security Institute cyber evaluation where safeguards were intentionally lowered. OpenAI’s GPT-5.6-Sol was also involved in similar events. These incidents highlight growing concerns about the potential misuse of advanced AI.