PASCAL - Pattern Analysis, Statistical Modelling and Computational Learning

Speaker Attribution in Cabinet Protocols
Josef Ruppenhofer, Caroline Sporleder and Fabian Shirokov
In: LREC 2010, 19-21 May 2010, Valetta, Malta.


Historical cabinet protocols are a useful resource which enable historians to identify the opinions expressed by politicians on different subjects and at different points of time. While cabinet protocols are often available in digitized form, so far the only method to access their information content is by keyword-based search, which often returns sub-optimal results. We present a method for enriching German cabinet protocols with information about the originators of statements. This requires automatic speaker attribution. Unlike many other approaches, our method can also deal with cases in which the speaker is not explicitly identified in the sentence itself. Such cases are very common in our domain. To avoid costly manual annotation of training data, we design a rule-based system which exploits morphosyntactic cues. We show that such a system obtains good results, especially with respect to recall which is particularly important for information access.

EPrint Type:Conference or Workshop Item (Paper)
Project Keyword:Project Keyword UNSPECIFIED
Subjects:Natural Language Processing
Information Retrieval & Textual Information Access
ID Code:7124
Deposited By:Caroline Sporleder
Deposited On:04 March 2011