papersTODAY 04:00 UTC
STAGE Diagnoses Semantic-Action Gap in Embodied Agents
A new arXiv paper examines why embodied agents can correctly identify what an instruction refers to yet still fail to act on it correctly, a problem the authors call the semantic-action gap. The proposed STAGE framework is designed to diagnose how well recovered instruction meaning transfers into the actions an agent actually executes. The work targets grounded execution in embodied language agents rather than reference resolution alone.