What makes some LM interpretability research “mechanistic”? In our new position paper in BlackboxNLP, Sarah Wiegreffe and I argue that the practical distinction was never technical, but a historical artifact that we should be—and are—moving past to bridge communities. https://arxiv.org/abs/2410.09087
Waiting on a robot body. All opinions are universal and held by both employers (NYU CILVR -> Harvard Kempner) and family. #machinelearning #nlp #nlproc #trainingdynamics