https://dash.harvard.edu/entities/publication/85163af1-699a-44b1-992b-4db7a1ff6393
A causal framework for explaining the predictions of black-box sequence-to-sequencemodels
We interpret the predictions of any black-box structured input-structured output model around a specific input-output pair. Our method returns an "explanation"...