Two MoTacon attendees are on the left. The MoTaacon logo is in the center, and to the right a prompt to Get Your Ticket.

Low-resource Languages

Low-resource Languages image
Languages that are significantly underrepresented in computational datasets and NLP research, typically because the volume of digitised text available for training is small relative to dominant languages such as English.

So what? Over 90% of the world's languages are classified as low-resource; this imbalance means AI models perform unevenly across linguistic and cultural contexts, producing representational injustice for speakers of those languages.

Example: A multilingual model trained predominantly on English-language web data may perform well for English speakers but generate unreliable or culturally inappropriate outputs for speakers of languages with limited digital corpora.
Explore MoT
Influence, from the other side of the table image
What I learned about influence by becoming a stakeholder
How to test GenAI Agents (by building one) image
Turn AI curiosity into engineering confidence.
This Week in Quality image
Debrief the week in Quality via a community radio show hosted by Simon Tomes and members of the community
Subscribe to our newsletter