AI & ML interests
We build tools to understand how models change during training, identify where regressions and unwanted behaviors emerge, localize meaningful changes within the model, and correct or remove learned behavior without full retraining.Our work spans training dynamics, model interpretability, machine unlearning, training-free model optimization, and AI governance.
Recent Activity
View all activity
Articles
authentrics/pii-removal-medical-llama32
Text Generation • Updated
authentrics/mnist-badnets-repaired
Image Classification • Updated
authentrics/nemotron-training-dynamics
Text Generation • Updated
authentrics/mnist-badnets-poisoned
Image Classification • Updated • 1
authentrics/mixtral-expert-routing
Text Generation • Updated
authentrics/ztom-gemma3-1b-recovery
Text Generation • Updated • 1
authentrics/gpt-oss-03
13.3M • Updated • 9
authentrics/gpt-oss-02
13.3M • Updated • 10
authentrics/gpt-oss-01
13.3M • Updated • 8
authentrics/gemma3-03
0.4B • Updated • 12
authentrics/gemma3-02
0.4B • Updated • 12
authentrics/gemma3-01
0.4B • Updated • 10
authentrics/mnist-03
Updated
authentrics/mnist-02
Updated
authentrics/mnist-01
Updated
authentrics/simple-dense-03
Updated
authentrics/simple-dense-02
Updated
authentrics/simple-dense-01
Updated
authentrics/toy-moe
2.36M • Updated • 7
authentrics/toy-gemma3
4.24M • Updated • 7
authentrics/toy_llm
Text Generation • 16.5M • Updated • 119
authentrics/medical-chatbot-pii-removed
Text Generation • Updated
authentrics/medical-chatbot-pii-overtrained
Text Generation • Updated
authentrics/medical-chatbot-pii-often
Text Generation • Updated