A General, Sound and Efficient Natural Language Parsing Algorithm based on Syntactic Constraints PropagationThis paper presents a new context-free parsing algorithm based on a bidirectional strictly horizontal strategy which incorporates strong top-down predictions (derivations and adjacencies). From a…Jose F. Quesada·Jan 26, 1998SaveLearn
Modularity in inductively-learned word pronunciation systemsIn leading morpho-phonological theories and state-of-the-art text-to-speech systems it is assumed that word pronunciation cannot be learned or performed without in-between analyses at several…Antal van den Bosch, Ton Weijters, Walter Daelemans·Jan 26, 1998SaveLearn
Do not forget: Full memory in memory-based learning of word pronunciationMemory-based learning, keeping full memory of learning material, appears a viable approach to learning NLP tasks, and is often superior in generalisation accuracy to eager learning approaches that…Antal van den Bosch, Walter Daelemans·Jan 26, 1998SaveLearn
Hierarchical Non-Emitting Markov ModelsWe describe a simple variant of the interpolated Markov model with non-emitting state transitions and prove that it is strictly more powerful than any Markov model. More importantly, the non-emitting…Eric Sven Ristad, Robert G. Thomas·Jan 20, 1998SaveLearn
Identifying Discourse Markers in Spoken DialogIn this paper, we present a method for identifying discourse marker usage in spontaneous speech based on machine learning. Discourse markers are denoted by special POS tags, and thus the process of…Peter A. Heeman, Donna Byron, James F. Allen·Jan 17, 1998SaveLearn
Orthographic Structuring of Human Speech and Texts: Linguistic Application of Recurrence Quantification AnalysisA methodology based upon recurrence quantification analysis is proposed for the study of orthographic structure of written texts. Five different orthographic data sets (20th century Italian poems,…F. Orsucci, K. Walter, A. Giuliani et al.·Dec 24, 1997SaveLearn
Speech Repairs, Intonational Boundaries and Discourse Markers: Modeling Speakers' Utterances in Spoken DialogIn this thesis, we present a statistical language model for resolving speech repairs, intonational boundaries and discourse markers. Rather than finding the best word interpretation for an acoustic…Peter A. Heeman·Dec 23, 1997SaveLearn
What is word sense disambiguation good for?Word sense disambiguation has developed as a sub-area of natural language processing, as if, like parsing, it was a well-defined task which was a pre-requisite to a wide range of…Adam Kilgarriff·Dec 23, 1997SaveLearn
Foreground and Background Lexicons and Word Sense Disambiguation for Information ExtractionLexicon acquisition from machine-readable dictionaries and corpora is currently a dynamic field of research, yet it is often not clear how lexical information so acquired can be used, or how it…Adam Kilgarriff·Dec 23, 1997SaveLearn
"I don't believe in word senses"Word sense disambiguation assumes word senses. Within the lexicography and linguistics literature, they are known to be very slippery entities. The paper looks at problems with existing accounts of…Adam Kilgarriff·Dec 23, 1997SaveLearn
Topic Graph Generation for Query Navigation: Use of Frequency Classes for Topic ExtractionTo make an interactive guidance mechanism for document retrieval systems, we developed a user-interface which presents users the visualized map of topics at each stage of retrieval process. Topic…Yoshiki Niwa, Shingo Nishioka, Makoto Iwayama et al.·Dec 12, 1997SaveLearn
Machine Learning of User Profiles: Representational IssuesAs more information becomes available electronically, tools for finding information of interest to users becomes increasingly important. The goal of the research described here is to build a system…Eric Bloedorn, Inderjeet Mani, T. Richard MacMillan·Dec 11, 1997SaveLearn
Multi-document Summarization by Graph Search and MatchingWe describe a new method for summarizing similarities and differences in a pair of related documents using a graph representation for text. Concepts denoted by words, phrases, and proper names in the…Inderjeet Mani, Eric Bloedorn·Dec 10, 1997SaveLearn
Context as a Spurious ConceptI take issue with AI formalizations of context, primarily the formalization by McCarthy and Buvac, that regard context as an undefined primitive whose formalization can be the same in many different…Graeme Hirst·Dec 9, 1997SaveLearn
Applying Explanation-based Learning to Control and Speeding-up Natural Language GenerationThis paper presents a method for the automatic extraction of subgrammars to control and speeding-up natural language generation NLG. The method is based on explanation-based learning (EBL). The main…Guenter Neumann·Dec 8, 1997SaveLearn
Type-driven semantic interpretation and feature dependencies in R-LFGOnce one has enriched LFG's formal machinery with the linear logic mechanisms needed for semantic interpretation as proposed by Dalrymple et. al., it is natural to ask whether these make any…Mark Johnson·Nov 21, 1997SaveLearn
The effect of alternative tree representations on tree bank grammarsThe performance of PCFGs estimated from tree banks is sensitive to the particular way in which linguistic constructions are represented as trees in the tree bank. This paper presents a theoretical…Mark Johnson·Nov 21, 1997SaveLearn
Features as Resources in R-LFGThis paper introduces a non-unification-based version of LFG called R-LFG (Resource-based Lexical Functional Grammar), which combines elements from both LFG and Linear Logic. The paper argues that a…Mark Johnson·Nov 20, 1997SaveLearn
Proof Nets and the Complexity of Processing Center-Embedded ConstructionsThis paper shows how proof nets can be used to formalize the notion of ``incomplete dependency'' used in psycholinguistic theories of the unacceptability of center-embedded constructions.…Mark Johnson·Nov 20, 1997SaveLearn
Application-driven automatic subgrammar extractionThe space and run-time requirements of broad coverage grammars appear for many applications unreasonably large in relation to the relative simplicity of the task at hand. On the other hand,…Renate Henschel, John A. Bateman·Nov 19, 1997SaveLearn
Towards an Improved Performance Measure for Language ModelsIn this paper a first attempt at deriving an improved performance measure for language models, the probability ratio measure (PRM) is described. In a proof of concept experiment, it is shown that PRM…Joerg P. Ueberla·Nov 19, 1997SaveLearn
On the use of expectations for detecting and repairing human-machine miscommunicationIn this paper I describe how miscommunication problems are dealt with in the spoken language system DIALOGOS. The dialogue module of the system exploits dialogic expectations in a twofold way: to…Morena Danieli·Nov 19, 1997SaveLearn
Language Modelling For Task-Oriented DomainsThis paper is focused on the language modelling for task-oriented domains and presents an accurate analysis of the utterances acquired by the Dialogos spoken dialogue system. Dialogos allows access…Cosmin Popovici, Paolo Baggia·Nov 19, 1997SaveLearn
Contextual Information and Specific Language Models for Spoken Language UnderstandingIn this paper we explain how contextual expectations are generated and used in the task-oriented spoken language understanding system Dialogos. The hard task of recognizing spontaneous speech on the…Paolo Baggia, Morena Danieli, Elisabetta Gerbino et al.·Nov 19, 1997SaveLearn
Some apparently disjoint aims and requirements for grammar development environments: the case of natural language generationGrammar development environments (GDE's) for analysis and for generation have not yet come together. Despite the fact that analysis-oriented GDE's (such as ALEP) may include some possibility…John A. Bateman·Nov 19, 1997SaveLearn