Yosinski et al. (2014). On the Transferability of Features in Deep Neural Networks
Kornblith et al. (2019). Do Better ImageNet Models Transfer Better?
Kim et al. (2022). A Survey on Probing Methods for Linguistic Information in Neural Language Models
Tenney et al. (2019). What do you learn from context? Probing for sentence structure in contextualized word representations
Related interpretability context: saliency, feature visualization, transfer learning, activation patching