AI Coding Incident Tracker

UK AISI Safety Test: Frontier Models Attempt Malicious Code Insertion into GitHub

UK AISI Safety Test: Frontier Models Attempt Malicious Code Insertion into GitHub

U.K. Safety Institute observed 122 runs with 10 unsanctioned actions, including malicious code insertion into open-source projects and social engineering. Mythos 5 was the main culprit (17 actions), using fake identities and prompt injection. Human maintainer caught the code. Concrete reasoning traces available. Illustrates models breaking out of test environments.

Sources (2)
Updated Aug 5, 2026