Automating the acquisition of bilingual terminology
Published 1 January 1993Open access
Pim van der Eijk
Citations68
Generate an AI Snapshot to get a quick, structured summary of this paper.
Study Snapshot
ObjectiveStudy objective
MethodsResearch methodology
PopulationPopulation studied
Sample sizeSample sizes
OutcomesStudy outcomes here
ResultsStudy results comes here
LimitationsResearch study limitations comes here
A concise AI-generated summary of the paper will appear here once you click Generate AI Snapshot.
TL;DR
Experimental results for a number of methods, which operate on corpora of previously translated texts, which compile bilingual lists of terminological expressions as automatically as possible are discussed.
Abstract
As the acquisition problem of bilingual lists of terminological expressions is formidable, it is worthwhile to investigate methods to compile such lists as automatically as possible. In this paper we discuss experimental results for a number of methods, which operate on corpora of previously translated texts.
Keywords
Computer Science
Word association norms, mutual information, and lexicography
3,674 Citations1990Kenneth Church, Patrick Hanks
A statistical approach to machine translation
1,699 Citations1990Peter F. Brown, John Cocke +6 more
The application of the statistical approach to translation from French to English and preliminary results are described and the results are given.
A stochastic parts program and noun phrase parser for unrestricted text
974 Citations1988Kenneth Church
A program that tags each word in an input sentence with the most likely part of speech has been written and performance is encouraging; a 400-word sample is presented and is judged to be 99.5% correct.
Aligning sentences in parallel corpora
487 Citations1991Peter F. Brown, Jennifer C. Lai +1 more
This paper describes a statistical technique for aligning sentences with their translations in two parallel corpora and shows that even without the benefit of anchor points the correlation between the lengths of aligned sentences is strong enough that it should be expected to achieve an accuracy of between 96% and 97%.
Identifying word correspondence in parallel texts
297 Citations1991William A. Gale, Kenneth Church
Researchers in both machine translation and bilingual lexicography have recently become interested in studying parallel texts, bodies of text which are available in multiple languages and who outline a self-organizing method for using these parallel texts to build a machine translation system.
A program for aligning sentences in bilingual corpora
258 Citations1991William A. Gale, Kenneth Church
A method for aligning sentences in parallel texts, such as the Canadian Hansards, based on a simple statistical model of character lengths is described, developed and tested on a small trilingual sample of Swiss economic reports.
Using Bilingual Materials to Develop Word Sense Disambiguation Methods
103 Citations2005William A. Gale, Kenneth Church +1 more
The methodology required to address the classic problem of word sense disambiguation is focused on, and there will be less to say about the details required for practical application of this methodology.
