David Bau
(he/him/his)
Assistant Professor
Research interests
- Machine learning
- Computer vision
- Artificial intelligence
- Natural language processing
- Human–computer interaction
Education
- PhD in Computer Science, Massachusetts Institute of Technology
- MS in Computer Science, Cornell University
- AB in Mathematics, Harvard University
Biography
David Bau is an assistant professor in the Khoury College of Computer Sciences at Northeastern University, based in Boston.
Bau's research focuses on human-computer interaction and machine learning. Before joining Northeastern, he worked as a software engineer at Google, BEA, and Crossgain. He has been published in journals such as CVPR, NeurIPS, ICCV, ECCV, and SIGGRAPH.
Outside of research, Bau enjoys astronomy and puzzle collecting.
Recent publications
-
SliderSpace: Decomposing the Visual Capabilities of Diffusion Models
Citation: Rohit Gandikota, Zongze Wu , Richard Zhang , David Bau, Eli Shechtman, Nicholas I. Kolkin. (2025). SliderSpace: Decomposing the Visual Capabilities of Diffusion Models ICCV, 15994-16003. https://doi.org/10.1109/ICCV51701.2025.01484 -
In-Context Algebra
Citation: Eric Todd, Jannik Brinkmann, Rohit Gandikota, David Bau. (2026). In-Context Algebra ICLR. https://proceedings.iclr.cc/paper_files/paper/2026/hash/82aec8518602748540a42b783468c94d-Abstract-Conference.html -
LLMs Process Lists With General Filter Heads
Citation: Arnab Sen Sharma, Giordano Rogers, Natalie Shapira, David Bau. (2026). LLMs Process Lists With General Filter Heads ICLR. https://proceedings.iclr.cc/paper_files/paper/2026/hash/6732332e3c4155da82cb9d4970e88133-Abstract-Conference.html -
Agents of Chaos
Citation: Natalie Shapira, Chris Wendler, Avery Yen, Gabriele Sarti, Koyena Pal, Olivia Floody, Adam Belfki, Alexander R. Loftus, Aditya Ratan Jannali, Nikhil Prakash, Jasmine Cui, Giordano Rogers, Jannik Brinkmann, Can Rager, Amir Zur, Michael Ripa, Aruna Sankaranarayanan, David Atkinson, Rohit Gandikota, Jaden Fiotto-Kaufman, EunJeong Hwang, Hadas Orgad, P. Sam Sahil, Negev Taglicht, Tomer Shabtay, Atai Ambus, Nitay Alon, Shiri Oron, Ayelet Gordon-Tapiero, Yotam Kaplan, Vered Shwartz, Tamar Rott Shaham, Christoph Riedl, Reuth Mirsky, Maarten Sap, David Manheim, Tomer Ullman, David Bau. (2026). Agents of Chaos CoRR, abs/2602.20021. https://doi.org/10.48550/arXiv.2602.20021 -
Machine Unlearning Doesn’t Do What You Think: Lessons for Generative AI Policy and Research
Citation: A. Feder Cooper, Christopher A. Choquette-Choo, Miranda Bogen, Kevin Klyman, Matthew Jagielski, Katja Filippova, Ken Liu, Alexandra Chouldechova, Jamie Hayes, Yangsibo Huang, Eleni Triantafillou, Peter Kairouz, Nicole Mitchell, Niloofar Mireshghallah, Abigail Z. Jacobs, James Grimmelmann, Vitaly Shmatikov, Christopher De Sa, Ilia Shumailov, Andreas Terzis, Solon Barocas, Jennifer Wortman Vaughan, danah boyd, Yejin Choi , Sanmi Koyejo, Fernando A. Delgado, Percy Liang, Daniel E. Ho, Pamela Samuelson, Miles Brundage, David Bau, Seth Neel, Hanna M. Wallach, Amy Cyphert, Mark A. Lemley, Nicolas Papernot, Katherine Lee. (2025). Machine Unlearning Doesn't Do What You Think: Lessons for Generative AI Policy and Research NeurIPS. http://papers.nips.cc/paper_files/paper/2025/hash/8110ef9878e4642ad848ca583d37d7ca-Abstract-Position_Paper_Track.html -
Elucidating Mechanisms of Demographic Bias in LLMs for Healthcare
Citation: Hiba Ahsan, Arnab Sen Sharma, Silvio Amir, David Bau, Byron C. Wallace. (2025). Elucidating Mechanisms of Demographic Bias in LLMs for Healthcare EMNLP (Findings), 14614-14631. https://doi.org/10.18653/V1/2025.FINDINGS-EMNLP.789 -
Position-aware Automatic Circuit Discovery
Citation: Tal Haklay, Hadas Orgad, David Bau, Aaron Mueller, Yonatan Belinkov. (2025). Position-aware Automatic Circuit Discovery ACL (1), 2792-2817. https://aclanthology.org/2025.acl-long.141/ -
Language Models use Lookbacks to Track Beliefs
Citation: Nikhil Prakash, Natalie Shapira, Arnab Sen Sharma, Christoph Riedl, Yonatan Belinkov, Tamar Rott Shaham, David Bau, Atticus Geiger. (2025). Language Models use Lookbacks to Track Beliefs CoRR, abs/2505.14685. https://doi.org/10.48550/arXiv.2505.14685 -
MIB: A Mechanistic Interpretability Benchmark
Citation: Aaron Mueller, Atticus Geiger, Sarah Wiegreffe, Dana Arad, Iván Arcuschin, Adam Belfki, Yik Siu Chan, Jaden Fried Fiotto-Kaufman, Tal Haklay, Michael Hanna , Jing Huang, Rohan Gupta, Yaniv Nikankin, Hadas Orgad, Nikhil Prakash, Anja Reusch, Aruna Sankaranarayanan, Shun Shao, Alessandro Stolfo, Martin Tutek, Amir Zur, David Bau, Yonatan Belinkov. (2025). MIB: A Mechanistic Interpretability Benchmark ICML. https://openreview.net/forum?id=sSrOwve6vb -
NNsight and NDIF: Democratizing Access to Open-Weight Foundation Model Internals
Citation: Jaden Fried Fiotto-Kaufman, Alexander Russell Loftus, Eric Todd, Jannik Brinkmann, Koyena Pal, Dmitrii Troitskii, Michael Ripa, Adam Belfki, Can Rager, Caden Juang, Aaron Mueller, Samuel Marks, Arnab Sen Sharma, Francesca Lucchetti, Nikhil Prakash, Carla E. Brodley, Arjun Guha, Jonathan Bell , Byron C. Wallace, David Bau. (2025). NNsight and NDIF: Democratizing Access to Open-Weight Foundation Model Internals ICLR. https://openreview.net/forum?id=MxbEiFRf39 -
Sparse Feature Circuits: Discovering and Editing Interpretable Causal Graphs in Language Models
Citation: Samuel Marks, Can Rager, Eric J. Michaud, Yonatan Belinkov, David Bau, Aaron Mueller. (2025). Sparse Feature Circuits: Discovering and Editing Interpretable Causal Graphs in Language Models ICLR. https://openreview.net/forum?id=I4e82CIDxv -
Token Erasure as a Footprint of Implicit Vocabulary Items in LLMs
Citation: Sheridan Feucht, David Atkinson, Byron C. Wallace, David Bau. (2024). Token Erasure as a Footprint of Implicit Vocabulary Items in LLMs EMNLP, 9727-9739. https://aclanthology.org/2024.emnlp-main.543 -
Linearity of Relation Decoding in Transformer Language Models
Citation: Evan Hernandez, Arnab Sen Sharma, Tal Haklay, Kevin Meng, Martin Wattenberg, Jacob Andreas, Yonatan Belinkov, David Bau. (2024). Linearity of Relation Decoding in Transformer Language Models ICLR. https://openreview.net/forum?id=w7LU2s14kE -
Function Vectors in Large Language Models
Citation: Eric Todd, Millicent L. Li, Arnab Sen Sharma, Aaron Mueller, Byron C. Wallace, David Bau. (2024). Function Vectors in Large Language Models ICLR. https://openreview.net/forum?id=AwyxtyMwaG -
Fine-Tuning Enhances Existing Mechanisms: A Case Study on Entity Tracking
Citation: Nikhil Prakash, Tamar Rott Shaham, Tal Haklay, Yonatan Belinkov, David Bau. (2024). Fine-Tuning Enhances Existing Mechanisms: A Case Study on Entity Tracking ICLR. https://openreview.net/forum?id=8sKcAWOf2D -
Erasing Concepts from Diffusion Models
Citation: Rohit Gandikota, Joanna Materzynska, Jaden Fiotto-Kaufman, David Bau. (2023). Erasing Concepts from Diffusion Models ICCV, 2426-2436. https://doi.org/10.1109/ICCV51070.2023.00230 -
Future Lens: Anticipating Subsequent Tokens from a Single Hidden State
Citation: Pal, K., Sun, J., Yuan, A., Wallace, B.C., & Bau, D. (2023). Future Lens: Anticipating Subsequent Tokens from a Single Hidden State. ArXiv, abs/2311.04897. -
FIND: A Function Description Benchmark for Evaluating Interpretability Methods
Citation: Sarah Schwettmann, Tamar Rott Shaham, Joanna Materzynska, Neil Chowdhury, Shuang Li, Jacob Andreas, David Bau, Antonio Torralba . (2023). FIND: A Function Description Benchmark for Evaluating Interpretability Methods NeurIPS. http://papers.nips.cc/paper_files/paper/2023/hash/ef0164c1112f56246224af540857348f-Abstract-Datasets_and_Benchmarks.html -
Emergent World Representations: Exploring a Sequence Model Trained on a Synthetic Task
Citation: Kenneth Li , Aspen K. Hopkins, David Bau, Fernanda B. Viégas, Hanspeter Pfister, Martin Wattenberg. (2023). Emergent World Representations: Exploring a Sequence Model Trained on a Synthetic Task ICLR. https://openreview.net/pdf?id=DeG07_TcZvT -
Mass-Editing Memory in a Transformer
Citation: Kevin Meng, Arnab Sen Sharma, Alex J. Andonian, Yonatan Belinkov, David Bau. (2023). Mass-Editing Memory in a Transformer ICLR. https://openreview.net/pdf?id=MkbcAHIYgyS -
Toward a Visual Concept Vocabulary for GAN Latent Space
Citation: Sarah Schwettmann, Evan Hernandez, David Bau, Samuel Klein, Jacob Andreas, Antonio Torralba . (2021). Toward a Visual Concept Vocabulary for GAN Latent Space ICCV, 6784-6792. https://doi.org/10.1109/ICCV48922.2021.00673 -
Locating and Editing Factual Associations in GPT
Citation: Meng, K., Bau, D., Andonian, A., & Belinkov, Y. (2022). Locating and Editing Factual Associations in GPT. Neural Information Processing Systems. -
Disentangling visual and written concepts in CLIP
Citation: Joanna Materzynska, Antonio Torralba , David Bau. (2022). Disentangling visual and written concepts in CLIP CVPR, 16389-16398. https://doi.org/10.1109/CVPR52688.2022.01592 -
Sketch Your Own GAN
Citation: Sheng-Yu Wang, David Bau, and Jun-Yan Zhu. Sketch Your Own GAN. Proceedings of the IEEE/CVF International Conference on Computer Vision. (ICCV 2021)