A Researcher Poisoned an Open-Weight AI Model for Under $100
TL;DR
- Katie Paxton-Fear, cybersecurity lecturer at Manchester Metropolitan University and security advocate at Semgrep, installed a backdoor in an open-weight coding model in about one hour for under $100. - She used just ten training examples to make the model produce code vulnerable to remote code execution (CWE-78)