Evaluating and Improving Multilinguality in Large Language Models

I'm interested in making LLMs better at multilinguality. There are several dimensions to this problem:

Here's a schema of my work categorised along the above dimensions! These are first author works, except for [4] and [9], in which I mentored students as the last author. [8] was a large team effort at Meta with 9 core contributors including me, for which I led the component on post-training dataset design.

Filter by language type:
Comprehension
Generation
Evaluation
Improving
Language Type
Techniques
a Data Strategies
b Dataset Design
c Modeling / Training
d Interpretability
c Theory