Multilingual Machine Translation in the Indian Subcontinent: Approaches, Challenges, and Cross-Lingual Paradigms
Keywords:
Machine Translation, Indian Languages, Computational Paninian Grammar, Structural Isomorphism, Phrase-Based SMT, Byte-Pair Encoding.Abstract
Machine Translation (MT) in the Indian subcontinent presents unique computational challenges due tohigh morphological complexity, flexible word order, and severe resource asymmetry among IndoAryan and Dravidian language families. This survey provides a comprehensive analysis of 51 translation architectures spanning three decades of Machine Translation
References
[1] R. M. K. Sinha, "Machine translation in India: A brief survey," Journal of Computer Science and Technology, vol. 5, no. 2, pp. 1111–1116, 2010.
[2] H. Choudhary, A. K. Pathak, and R. R. Shah, "Neural machine translation for English-Tamil," in Proc. 3rd Conf. Machine Translation (WMT), Brussels, Belgium, 2018, pp. 783–788.


