Evaluating parts-of-speech taggers for use in a text-to-scene conversion system


Autoria(s): Glass, Kevin; Bangay, Shaun
Contribuinte(s)

Bishop, Judith

Kourie, Derrick

Data(s)

01/01/2005

Resumo

This paper presents parts-of-speech tagging as a first step towards an autonomous text-to-scene conversion system. It categorizes some freely available taggers, according to the techniques used by each in order to automatically identify word-classes. In addition, the performance of each identified tagger is verified experimentally. The SUSANNE corpus is used for testing and reveals the complexity of working with different tagsets, resulting in substantially lower accuracies in our tests than in those reported by the developers of each tagger. The taggers are then grouped to form a voting system to attempt to raise accuracies, but in no cases do the combined results improve upon the individual accuracies. Additionally a new metric, agreement, is tentatively proposed as an indication of confidence in the output of a group of taggers where such output cannot be validated.<br />

Identificador

http://hdl.handle.net/10536/DRO/DU:30039203

Idioma(s)

eng

Publicador

South African Institute for Computer Scientists and Information Technologists

Relação

http://dro.deakin.edu.au/eserv/DU:30039203/bangay-evaluatingparts-2005.pdf

http://dl.acm.org/citation.cfm?id=1145678&CFID=54220471&CFTOKEN=33121831

Direitos

2005, SAICSIT

Palavras-Chave #corpora #parts-of-speech tagging
Tipo

Conference Paper