Alina Shabaeva
2026
LLMs Out-of-the-Box Do Not Generate Context-Appropriate Word Order in Russian
Alina Shabaeva | John Frederick Bailyn | Owen Rambow
Proceedings of the 27th Annual Meeting of the Special Interest Group on Discourse and Dialogue
Alina Shabaeva | John Frederick Bailyn | Owen Rambow
Proceedings of the 27th Annual Meeting of the Special Interest Group on Discourse and Dialogue
In this paper, we examine whether LLMs can generate context-appropriate word order in Russian. Using a new corpus of movie scripts in Russian, with a particular focus on dialogues, we investigate how various semantic, morphosyntactic, and discourse features influence word order choice. We show that traditional machine learning can use these features to model different word order fairly well. We also examine whether LLM dialogue partners can generate context-appropriate word order in Russian. With both zero- and few-shot prompting, LLMs fail to generate word orders beyond the default subject-verb-object. Instead, in order to generate context-appropriate word orders, LLMs need to be fine-tuned or given explicit word order suggestions from a traditional machine learning method.