Developing language resources for English machine translation

Thumbnail Image

Date

2008-08

Journal Title

Journal ISSN

Volume Title

Publisher

BRAC University

Abstract

We developed English-Bangla parallel corpora for statistical machine translation. By hand we tagged 20,000 words of our Bangla corpus according to their particular part of speeches. In our work we also suggested a method for identifying word correspondence in parallel English-Bangla text using a translation model based on part of speech and n-gram model.

Description

Cataloged from PDF version of thesis report.
Includes bibliographical references (page 93).
This thesis report is submitted in partial fulfillment of the requirements for the degree of Bachelor of Science in Computer Science and Engineering, 2008.

Keywords

Computer science and engineering

Citation

Endorsement

Review

Supplemented By

Referenced By