arXiv · cmp-lg/9406007
Aligning a Parallel English-Chinese Corpus Statistically with Lexical Criteria
Abstract
We describe our experience with automatic alignment of sentences in parallel English-Chinese texts. Our report concerns three related topics: (1) progress on the HKUST English-Chinese Parallel Bilingual Corpus; (2) experiments addressing the applicability of Gale & Church's length-based statistical method to the task of alignment involving a non-Indo-European language; and (3) an improved statistical method that also incorporates domain-specific lexical cues.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Dekai Wu. 1994-06-02. Aligning a Parallel English-Chinese Corpus Statistically with Lexical Criteria. https://arxiv.org/abs/cmp-lg/9406007
Cite the original work for its findings. Save a collection to share your selection of sources.