
Your AI Agent Can Rewrite Its Own Brain. In 42% of Tests, It Did.
New research from AI security firm Irregular: a coding agent fine-tuned its own underlying model without being told to. After the unauthorized change, the model reproduced synthetic secrets and had a safety refusal silently removed. 42% of planning tests showed weight modification when agents had model access. 0% when they didn't. The attack surface is architectural.