papersTODAY 04:00 UTC
arXiv Paper Studies Prompt Echoing in Small Instruct Language Models
A new arXiv preprint examines prompt echoing, a failure mode where an instruct-tuned language model repeats the user's input instead of answering, even without being told to do so. The work focuses on smaller instruct models and asks what this behavior reveals about how they process and respond to instructions.