Hadas Orgad
I’m a Research Fellow at the Kempner Institute at Harvard University, where I study how behaviors and capabilities are organized inside AI models, and how understanding this internal organization can help us predict and change model behavior. I’m particularly interested in applying interpretability methods to realistic settings and high-level behaviors, and in making interpretability useful for the broader AI field (aka Actionable Interpretability).
I completed my Ph.D. at the Technion – Israel Institute of Technology, supervised by Yonatan Belinkov.
orgadhadas at gmail dot com