James A. Michaelov, Carmen Amo Alonso, Tyler A. Chang, Roger P. LEVY · arXiv (Cornell University) 2026 · 2026
DOI: 10.48550/arxiv.2609.16967
Counts differ because each database indexes a different set of publications. We treat OpenAlex as the canonical count; Google Scholar is not shown (no API, and crawling it violates its ToS).
Ensuring that multilingual language models generate coherent text in a specific target language is a major issue in multilingual language modeling. We develop an optimal control method for target-language text generation as well as a framework for evaluating the quality of generated text in terms of language adherence, linguistic coherence, and semantic coherence. We find that the proposed method performs at least as well as the prominent difference-in-means activation steering method for the majority of models tested, with substantially less hyperparameter tuning required.
No comments yet — start the discussion below.