papersSEP 10 04:00 UTC
Study: Multilingual Data Mixing Helps LLMs Reason in Non-English Languages
A new arXiv paper tackles the tendency of reasoning language models to think in English even when prompted in other languages, which limits access for non-English speakers. The authors show that how training data is mixed across languages is key to getting models to generalize and carry out reasoning in the user's language itself. The work offers a path toward making advanced reasoning capabilities usable beyond English.