DCU at the NTCIR-9 spokendoc passage retrieval task
Eskevich, Maria and Jones, Gareth J.F. (2011) DCU at the NTCIR-9 spokendoc passage retrieval task. In: The 9th NTCIR Workshop Meeting , 6-9 Dec 2011, Tokyo, Japan. ISBN 978-4-86049-056-0
Full text available as:
We describe details of our runs and the results obtained for the "IR for Spoken Documents (SpokenDoc) Task" at NTCIR-9. The focus of our participation in this task was the investigation of the use of segmentation methods to divide the manual and ASR transcripts into topically coherent segments. The underlying assumption of this approach is that these segments will capture passages in the transcript relevant to the query. Our experiments investigate the use of two lexical coherence based segmentation algorithms (Text-Tiling, C99). These are run on the provided manual and ASR transcripts, and the ASR transcript with stop words removed. Evaluation of the results shows that TextTiling consistently performs better than C99 both in segmenting the data into retrieval units as evaluated using the centre located relevant information metric and in having higher content precision in each automatically created segment.
Archive Staff Only: edit this record