Analysis of the Arabic Broken Plural and DiminutiveThis paper demonstrates how the challenging problem of the Arabic broken plural and diminutive can be handled under a multi-tape two-level model, an extension to two-level morphology.George A. Kiraz·Dec 9, 1995SaveLearn
Using Information Content to Evaluate Semantic Similarity in a TaxonomyThis paper presents a new measure of semantic similarity in an IS-A taxonomy, based on the notion of information content. Experimental evaluation suggests that the measure performs encouragingly well…Philip Resnik·Nov 29, 1995SaveLearn
Disambiguating Noun Groupings with Respect to WordNet SensesWord groupings useful for language processing tasks are increasingly available, as thesauri appear on-line, and as distributional word clustering techniques improve. However, for many tasks, one is…Philip Resnik·Nov 29, 1995SaveLearn
Chart-driven Connectionist Categorial Parsing of Spoken KoreanWhile most of the speech and natural language systems which were developed for English and other Indo-European languages neglect the morphological processing and integrate speech and natural language…WonIl Lee, Geunbae Lee, Jong-Hyeok Lee·Nov 29, 1995SaveLearn
An investigation into the correlation of cue phrases, unfilled pauses and the structuring of spoken discourseExpectations about the correlation of cue phrases, the duration of unfilled pauses and the structuring of spoken discourse are framed in light of Grosz and Sidner's theory of discourse and are…Janet Cahn·Nov 22, 1995SaveLearn
The Effect of Resource Limits and Task Complexity on Collaborative Planning in DialogueThis paper shows how agents' choice in communicative action can be designed to mitigate the effect of their resource limits in the context of particular features of a collaborative planning task.…Marilyn A. Walker·Nov 15, 1995SaveLearn
Letting the Cat out of the Bag: Generation for Shake-and-Bake MTDescribes an algorithm for the generation phase of a Shake-and-Bake Machine Translation system. Since the problem is NP-complete, it is unlikely that the algorithm will be efficient in all cases, but…Chris Brew·Nov 13, 1995SaveLearn
Countability and Number in Japanese-to-English Machine TranslationThis paper presents a heuristic method that uses information in the Japanese text along with knowledge of English countability and number stored in transfer dictionaries to determine the countability…Francis Bond, Kentaro Ogura, Satoru Ikehara·Nov 3, 1995SaveLearn
Toward an MT System without Pre-Editing --- Effects of New Methods in ALT-J/E ---Recently, several types of Japanese-to-English machine translation systems have been developed, but all of them require an initial process of rewriting the original text into easily translatable…Satoru Ikehara, Satoshi Shirai, Akio Yokoo et al.·Oct 31, 1995SaveLearn
Automatic Identification of Support Verbs: A Step Towards a Definition of Semantic WeightCurrent definitions of notions of lexical density and semantic weight are based on the division of words into closed and open classes, and on intuition. This paper develops a computationally…Mark Dras·Oct 25, 1995SaveLearn
Incorporating Discourse Aspects in English -- Polish MT: Towards Robust ImplementationThe main aim of translation is an accurate transfer of meaning so that the result is not only grammatically and lexically correct but also communicatively adequate. This paper stresses the need for…Malgorzata E. Stys, Stefan S. Zemke·Oct 15, 1995SaveLearn
Developing and Evaluating a Probabilistic LR Parser of Part-of-Speech and Punctuation LabelsWe describe an approach to robust domain-independent syntactic parsing of unrestricted naturally-occurring (English) input. The technique involves parsing sequences of part-of-speech and punctuation…Ted Briscoe, John Carroll·Oct 9, 1995SaveLearn
Disambiguating bilingual nominal entries against WordNetThis paper explores the acquisition of conceptual knowledge from bilingual dictionaries (French/English, Spanish/English and English/Spanish) using a pre-existing broad coverage Lexical Knowledge…German Rigau, Eneko Agirre·Oct 4, 1995SaveLearn
A Proposal for Word Sense Disambiguation using Conceptual DistanceThis paper presents a method for the resolution of lexical ambiguity and its automatic evaluation over the Brown Corpus. The method relies on the use of the wide-coverage noun taxonomy of WordNet and…Eneko Agirre, German Rigau·Oct 4, 1995SaveLearn
POS Tagging Using Relaxation LabellingRelaxation labelling is an optimization technique used in many fields to solve constraint satisfaction problems. The algorithm finds a combination of values for a set of variables such that satisfies…Lluis Padro·Oct 2, 1995SaveLearn
ParseTalk about Textual EllipsisA hybrid methodology for the resolution of text-level ellipsis is presented in this paper. It incorporates conceptual proximity criteria applied to ontologically well-engineered domain knowledge…Michael Strube, Udo Hahn·Sep 28, 1995SaveLearn
Using Chinese Text Processing Technique for the Processing of Sanskrit Based Indian Languages: Maximum Resource Utilization and Maximum CompatibilityChinese text processing systems are using Double Byte Coding , while almost all existing Sanskrit Based Indian Languages have been using Single Byte coding for text processing. Through observation,…Md Maruf Hasan·Sep 27, 1995SaveLearn
The Development and Migration of Concepts from Donor to Borrower Disciplines: Sublanguage Term Use in Hard & Soft SciencesAcademic disciplines, often divided into hard and soft sciences, may be understood as "donor disciplines" if they produce more concepts than they borrow from other disciplines, or…Robert M. Losee·Sep 13, 1995SaveLearn
Cluster Expansions and Iterative Scaling for Maximum Entropy Language ModelsThe maximum entropy method has recently been successfully introduced to a variety of natural language applications. In each of these applications, however, the power of the maximum entropy method is…John D. Lafferty, Bernhard Suhm·Sep 9, 1995SaveLearn
Conserving Fuel in Statistical Language Learning: Predicting Data RequirementsIn this paper I address the practical concern of predicting how much training data is sufficient for a statistical language learning system. First, I briefly review earlier results and show how these…Mark Lauer·Sep 7, 1995SaveLearn
How much is enough?: Data requirements for statistical NLPIn this paper I explore a number of issues in the analysis of data requirements for statistical NLP systems. A preliminary framework for viewing such systems is proposed and a sample of existing…Mark Lauer·Sep 7, 1995SaveLearn
A Natural Law of SuccessionConsider the problem of multinomial estimation. You are given an alphabet of k distinct symbols and are told that the i-th symbol occurred exactly ni times in the past. On the basis of this…Eric Sven Ristad·Aug 30, 1995SaveLearn
The Use of Knowledge Preconditions in Language ProcessingIf an agent does not possess the knowledge needed to perform an action, it may privately plan to obtain the required information on its own, or it may involve another agent in the planning process by…Karen E. Lochbaum·Aug 29, 1995SaveLearn
Heuristics and Parse RankingThere are currently two philosophies for building grammars and parsers -- Statistically induced grammars and Wide-coverage grammars. One way to combine the strengths of both approaches is to have a…B. Srinivas, Christine Doran, Seth Kulick·Aug 28, 1995SaveLearn
A Labelled Analytic Theorem Proving Environment for Categorial GrammarWe present a system for the investigation of computational properties of categorial grammar parsing based on a labelled analytic tableaux theorem prover. This proof method allows us to take a modular…Saturnino F. Luz-Filho, Patrick Sturt·Aug 15, 1995SaveLearn