papersTODAY 04:00 UTC
arXiv paper proposes scalable data attribution via influence matrix estimation
A new arXiv preprint addresses the computational cost of data attribution, which measures how individual training samples affect a model's behavior. The authors frame the problem around estimating the influence matrix at scale, with applications in data valuation, machine unlearning, and interpretability. The abstract highlights that scaling such methods has remained a longstanding obstacle.