papersTODAY 04:00 UTC
Study Uses Activation Patching to Trace How VLMs Read Bar Chart Values
A new arXiv paper examines how vision-language models arrive at exact values when reading vertical bar charts. The authors apply counterfactual activation patching to trace where and how chart evidence is combined across space and depth in the network. The work argues that correct answers alone do not reveal the underlying mechanisms models use.