Back to Search Start Over

Transliteration in Any Language with Surrogate Languages

Authors :
Mayhew, Stephen
Christodoulopoulos, Christos
Roth, Dan
Publication Year :
2016

Abstract

We introduce a method for transliteration generation that can produce transliterations in every language. Where previous results are only as multilingual as Wikipedia, we show how to use training data from Wikipedia as surrogate training for any language. Thus, the problem becomes one of ranking Wikipedia languages in order of suitability with respect to a target language. We introduce several task-specific methods for ranking languages, and show that our approach is comparable to the oracle ceiling, and even outperforms it in some cases.

Details

Database :
arXiv
Publication Type :
Report
Accession number :
edsarx.1609.04325
Document Type :
Working Paper