GILTIA
GILTIA
News | Press
People
Events
Publications
Contact
English
English
Español
Article
The YUA-ES Communicative Contexts Corpus: An Open Parallel Dataset of Everyday Yucatec Maya–Spanish Phrases
An openly licensed parallel corpus of 14,332 aligned phrase pairs in Yucatec Maya and Spanish, organized into 31 everyday communicative contexts — the first open, phrase-aligned corpus of its kind.
Alejandro Molina Villegas
,
María Elisa Chavarrea-Chim
,
Samuel Canul-Yah
,
Sary Lorena Hau-Ucán
Jun 17, 2026
PDF
Cite
DOI
View on Research Square
The YUA-ES Communicative Contexts Corpus: An Open Parallel Dataset of Everyday Yucatec Maya–Spanish Phrases
An openly licensed parallel corpus of 14,332 aligned phrase pairs in Yucatec Maya and Spanish, organized into 31 everyday communicative contexts — the first open, phrase-aligned corpus of its kind.
Alejandro Molina Villegas
,
María Elisa Chavarrea-Chim
,
Samuel Canul-Yah
,
Sary Lorena Hau-Ucán
PDF
Cite
DOI
View on Research Square
Generating a Culturally and Linguistically Adapted Word Similarity Benchmark for Yucatec Maya
Un benchmark de similitud de palabras cultural y lingüísticamente adaptado al maya yucateco, construido a partir de una Lista de Swadesh filtrada, que muestra que la elección del benchmark pesa más que el ajuste de hiperparámetros al evaluar word embeddings.
Alejandro Molina Villegas
,
Joel Suro-Villalobos
,
Jorge Reyes-Magaña
,
Silvia Fernández-Sabido
Sep 25, 2025
DOI
Ver en Inteligencia Artificial
Generating a Culturally and Linguistically Adapted Word Similarity Benchmark for Yucatec Maya
A culturally and linguistically adapted word similarity benchmark for Yucatec Maya, built from a filtered Swadesh List, showing that benchmark selection matters more than hyperparameter tuning when evaluating word embeddings.
Alejandro Molina Villegas
,
Joel Suro-Villalobos
,
Jorge Reyes-Magaña
,
Silvia Fernández-Sabido
DOI
View at Inteligencia Artificial
Cite
×