Experiences with the GTU grammar development environmentIn this paper we describe our experiences with a tool for the development and testing of natural language grammars called GTU (German: Grammatik-Testumgebumg; grammar test environment). GTU supports…Martin Volk, Dirk Richarz·Jul 21, 1997SaveLearn
Recognizing Referential Links: An Information Extraction PerspectiveWe present an efficient and robust reference resolution algorithm in an end-to-end state-of-the-art information extraction system, which must work with a considerably impoverished syntactic analysis…Megumi Kameyama·Jul 18, 1997SaveLearn
Stressed and Unstressed Pronouns: Complementary PreferencesI present a unified account of interpretation preferences of stressed and unstressed pronouns in discourse. The central intuition is the Complementary Preference Hypothesis that predicts the…Megumi Kameyama·Jul 18, 1997SaveLearn
Tailored Patient Information: Some Issues and QuestionsTailored patient information (TPI) systems are computer programs which produce personalised heath-information material for patients. TPI systems are of growing interest to the natural-language…Ehud Reiter, Liesl Osman·Jul 18, 1997SaveLearn
Finite State Transducers Approximating Hidden Markov ModelsThis paper describes the conversion of a Hidden Markov Model into a sequential transducer that closely approximates the behavior of the stochastic model. This transformation is especially…Andre Kempe·Jul 17, 1997SaveLearn
Intrasentential Centering: A Case StudyOne of the necessary extensions to the centering model is a mechanism to handle pronouns with intrasentential antecedents. Existing centering models deal only with discourses consisting of simple…Megumi Kameyama·Jul 16, 1997SaveLearn
Discourse Preferences in Dynamic LogicIn order to enrich dynamic semantic theories with a `pragmatic' capacity, we combine dynamic and nonmonotonic (preferential) logics in a modal logic setting. We extend a fragment of Van Benthem and…Jan Jaspars, Megumi Kameyama·Jul 16, 1997SaveLearn
A Flexible POS tagger Using an Automatically Acquired Language ModelWe present an algorithm that automatically learns context constraints using statistical decision trees. We then use the acquired constraints in a flexible POS tagger. The tagger is able to use…Lluis Marquez, Lluis Padro·Jul 11, 1997SaveLearn
Automatic Detection of Text GenreAs the text databases available to users become larger and more heterogeneous, genre becomes increasingly important for computational linguistics as a complement to topical and structural principles…Brett Kessler, Geoffrey Nunberg, Hinrich Schuetze·Jul 8, 1997SaveLearn
Reluctant Paraphrase: Textual Restructuring under an Optimisation ModelThis paper develops a computational model of paraphrase under which text modification is carried out reluctantly; that is, there are external constraints, such as length or readability, on an…Mark Dras·Jul 3, 1997SaveLearn
Learning Parse and Translation Decisions From Examples With Rich ContextWe propose a system for parsing and translating natural language that learns from examples and uses some background knowledge. As our parsing model we choose a deterministic shift-reduce type parser…Ulf Hermjakob·Jun 30, 1997SaveLearn
Efficient Construction of Underspecified Semantics under Massive AmbiguityWe investigate the problem of determining a compact underspecified semantical representation for sentences that may be highly ambiguous. Due to combinatorial explosion, the naive method of building…Jochen Doerre·Jun 26, 1997SaveLearn
Automatic Discovery of Non-Compositional Compounds in Parallel DataAutomatic segmentation of text into minimal content-bearing units is an unsolved problem even for languages like English. Spaces between words offer an easy first approximation, but this…I. Dan Melamed·Jun 24, 1997SaveLearn
A Word-to-Word Model of Translational EquivalenceMany multilingual NLP applications need to translate words between different languages, but cannot afford the computational expense of inducing or applying a full translation model. For these…I. Dan Melamed·Jun 24, 1997SaveLearn
A Portable Algorithm for Mapping Bitext CorrespondenceThe first step in most empirical work in multilingual NLP is to construct maps of the correspondence between texts and their translations ( bitext maps). The Smooth Injective Map Recognizer…I. Dan Melamed·Jun 24, 1997SaveLearn
A Lexicalist Approach to the Translation of Colloquial TextColloquial English (CE) as found in television programs or typical conversations is different than text found in technical manuals, newspapers and books. Phrases tend to be shorter and less…Fred Popowich, Davide Turcato, Olivier Laurens et al.·Jun 18, 1997SaveLearn
An Information Extraction Core System for Real World German Text ProcessingThis paper describes SMES, an information extraction core system for real world German text processing. The basic design criterion of the system is of providing a set of basic powerful, robust, and…G. Neumann, R. Backofen, J. Baur et al.·Jun 18, 1997SaveLearn
Three Generative, Lexicalised Models for Statistical ParsingIn this paper we first propose a new statistical parsing model, which is a generative model of lexicalised context-free grammar. We then extend the model to include a probabilistic treatment of both…Michael Collins·Jun 17, 1997SaveLearn
An Efficient Distribution of Labor in a Two Stage Robust Interpretation ProcessAlthough Minimum Distance Parsing (MDP) offers a theoretically attractive solution to the problem of extragrammaticality, it is often computationally infeasible in large scale practical applications.…Carolyn Penstien Rose', Alon Lavie·Jun 17, 1997SaveLearn
An Empirical Approach to Temporal Reference ResolutionThis paper presents the results of an empirical investigation of temporal reference resolution in scheduling dialogs. The algorithm adopted is primarily a linear-recency based approach that does not…Janyce Wiebe, Tom O'Hara, Kenneth McKeever et al.·Jun 16, 1997SaveLearn
A Model of Lexical Attraction and RepulsionThis paper introduces new methods based on exponential families for modeling the correlations between words in text and speech. While previous work assumed the effects of word co-occurrence…Doug Beeferman, Adam Berger, John Lafferty·Jun 13, 1997SaveLearn
Evaluating Competing Agent Strategies for a Voice Email AgentThis paper reports experimental results comparing a mixed-initiative to a system-initiative dialog strategy in the context of a personal voice email agent. To independently test the effects of dialog…Marilyn Walker, Donald Hindle, Jeanne Fromer et al.·Jun 13, 1997SaveLearn
Name Searching and Information RetrievalThe main application of name searching has been name matching in a database of names. This paper discusses a different application: improving information retrieval through name recognition. It…Paul Thompson, Christopher C. Dozier·Jun 12, 1997SaveLearn
Text Segmentation Using Exponential ModelsThis paper introduces a new statistical approach to partitioning text automatically into coherent segments. Our approach enlists both short-range and long-range language models to help it sniff out…Doug Beeferman, Adam Berger, John Lafferty·Jun 11, 1997SaveLearn
Determining Internal and External Indices for Chart GenerationThis paper presents a compilation procedure which determines internal and external indices for signs in a unification based grammar to be used in improving the computational efficiency of lexicalist…Arturo Trujillo·Jun 11, 1997SaveLearn