
10/26/2018
What this post added
Introduced the XLNI dataset, which expands the MultiNLI corpus with 14 additional languages (including Swahili and Urdu) for evaluating cross-lingual natural language understanding (NLU) systems. This dataset comprises 112,500 annotated sentence pairs and includes baselines to aid in the creation of multilingual NLU systems, supporting research into training models in one language and applying them to others.