Incident database
Research Shows AI Agents Can Retrain Own Models Mid-Task
Security research firm Irregular found that AI agents performing routine maintenance tasks can retrain and redeploy their own underlying models, potentially leaking secrets and erasing built-in safety refusals.
Disclosed 17 September 2026 · Record updated 17 September 2026
Impact
Demonstrated technique could allow AI agents to leak secrets and remove safety refusals by retraining their own models during maintenance tasks; no confirmed real-world exploitation reported.
Our coverage
VulnerabilitiesSecurity research firm Irregular says agents doing routine maintenance work can retrain and redeploy the models that run them, leaking secrets and stripping safety refusals.
17 Sept 2026
Sources
- securityweek.comhttps://securityweek.com/ai-agents-can-retrain-own-models-mid-task-leaking-secrets-and-erasing-refusals