CentroGeo: introduces the Maya–Spanish parallel corpus YUA-ES-CCC

CentroGeo Centro Público de Investigación, Facebook post, July 1, 2026

CentroGeo Centro Público de Investigación

We present the YUA-ES Communicative Contexts Corpus (YUA-ES-CCC), an open collection of 14,332 phrase pairs in Yucatec Maya and Spanish, organized into 31 everyday contexts: greetings, transportation, everyday conversation, family, food, and more.

More than 14,000 phrases in Yucatec Maya and Spanish, available as open data.

Maya–Spanish parallel corpus (YUA-ES-CCC) An open resource that drives research, education, and the development of language technologies for Indigenous languages.

Learn more at: cgeo.mx/#d35

Project developed in collaboration with Sedeculta Yucatán (Secretaría de Ciencias, Humanidades, Tecnología e Innovación).

The corpus is described in the publication The YUA-ES Communicative Contexts Corpus, available as a preprint on Research Square.

Post by CentroGeo Centro Público de Investigación on Facebook, July 1, 2026. Translated from the original Spanish for this archive.

Jaziel A. Carballo Tadeo
Jaziel A. Carballo Tadeo
PhD Candidate in Geospatial Information Sciences

PhD candidate at CentroGeo’s National Geointelligence Laboratory, researching the use of artificial intelligence for the preservation of Yucatec Maya.