Test-time training shown to boost in-context learning of nonlinear functions
A new arXiv paper examines test-time training (TTT), a method where selected model parameters are updated before each prediction so the model can adapt to test data. The authors note that while TTT has had empirical success, its theoretical basis is not well understood, and their analysis focuses on how it affects in-context learning of nonlinear functions.