Annotator disagreement in temporal laughter localization found to be structured, not random
A new paper studies how human labelers disagree about the precise onset and boundaries of laughter when annotating audio. The authors show that this disagreement follows systematic patterns rather than acting as random noise, challenging the common practice of scoring temporal laughter localization against a single reference annotation. They argue that evaluation protocols should instead account for the structured nature of annotator disagreement.