AI & ML interests
We build tools to understand how models change during training, identify where regressions and unwanted behaviors emerge, localize meaningful changes within the model, and correct or remove learned behavior without full retraining.Our work spans training dynamics, model interpretability, machine unlearning, training-free model optimization, and AI governance.
Recent Activity
View all activity