Robuta

https://arxiv.org/abs/2601.21996 [2601.21996] Mechanistic Data Attribution: Tracing the Training Origins of Interpretable LLM Units Abstract page for arXiv paper 2601.21996: Mechanistic Data Attribution: Tracing the Training Origins of Interpretable LLM Units data attributionthe traininginterpretable llmmechanistictracing