Assitant professor, MCA
Online published on 18 October, 2019.
In this research paper, we have worked upon the problem of “development of English-Punjabi parallel corpus and Unicode Character Mappingusing existing English-Punjabi machine transliteration system and using sentence alignment”. The alignment is based on the length and location based technique. We will use English-Punjabi machine transliteration system. These tasks are need to English-Punjabi parallel corpus and Unicode Mapping. Sentence alignment is useful for developing English-Punjabi parallel corpus and English-Punjabi dictionary. The accuracy is basically depending upon the complexity of the corpus and correspondence mapping, more the complexity less the accuracy. Complexity means how to distribution of sentence in the target file. If any of these categories 1: 1, 1: 2, 2: 1, 1: 3, 3: 1 sentences occur simultaneously in a paragraph. Our objective in this paper is to develop English-Punjabi parallel corpus and Unicode Mapping using latest and existing techniques and method with a high accuracy and time efficiency.