Igor Shilov
PhD student at Imperial College London working on AI security & privacy. Big Tech dropout.
about
Hi, I’m Igor!
I’m an AI security & privacy researcher in the final year of my PhD at Imperial College London’s AI Security & Privacy Lab, supervised by Yves-Alexandre de Montjoye.
I was part of the inaugural Anthropic AI Safety Fellowship Program, where I worked on capability removal, and a Research Fellow at Goodfire AI, where I studied mechanistic interpretability.
Before my PhD, I spent over a decade as a software engineer. Most recently, I was at Meta AI as a Research Engineer in Privacy Preserving ML group, where I led the development of Opacus, an open-source PyTorch library for differentially private model training. I also worked on StopNCII.org, a privacy-preserving platform that uses on-device perceptual hashing to combat non-consensual intimate image sharing.
My research interests include:
- 🏛️ Adversarial robustness in LLMs
- 🏛️ Memorization in LLMs
- 🏛️ Model internals
contact
I echo the standing invitation: I like getting email and I like talking to people. Please do reach out if you’re interested in discussing research, collaboration opportunities, or the future of privacy in AI!
I can also be found in Twitter.
selected publications
- ICML 2024Copyright Traps for Large Language ModelsIn Forty-first International Conference on Machine Learning (ICML), 2024
Press coverage in MIT Technology Review and Nature News.
selected projects
-
Opacus: User-Friendly Differential Privacy Library in PyTorchI was the lead developer and maintainer of Opacus, a PyTorch library to train ML models with Differential Privacy. With 1k+ stars on github and 300+ citations, the tool is helping to advance the state-of-the-art in privacy preserving ML both internally and for the wider community of researchers.
-
StopNCII.org: Stop Non-Consensual Intimate Image AbuseI have lead the team developing a privacy-preserving platform helping combat non-consensual intimate image sharing, a joint effort between Meta and a UK-based NGO running “Revenge Porn Helpline”. The platform takes advantage of on-device perceptual hashing to protect privacy. Since it’s launch, many major social media platforms has signed up as industry partners, inclusing Reddit, Snap, TilTok, OnlyFans and PornHub.
Selected press: NBC News, Bloomberg, Mashable, Cosmopolitan