Google DeepMind
I am a Research Scientist at Google DeepMind working on AGI readiness. My interests lie at the intersection of AI interpretability and AI responsibility. I study both black-box and mechanistic interpretability to bridge the gap between observable behaviors and their underlying neural mechanisms. My current focus is on emergent misbehavior in continual learning systems.
Previously, I worked at Meta on privacy mechanisms, federated learning, multimedia retrieval, and copyright protection.