I am an Associate Professor of EECS at MIT, where I run the Algorithmic Alignment Group in the Computer Science and Artificial Intelligence Laboratory (CSAIL). I work on AI alignment: developing methods to ensure that AI systems behave in accordance with the goals and values of their users and of society as a whole. My group's research spans interpreting and evaluating the internals of modern AI systems, learning robustly from human feedback, and designing the multi-agent and societal structures needed to oversee them. Our goal is to enable the safe, beneficial, and trustworthy deployment of AI in real-world settings.