Text categorization methods for automatic estimation of verbal intelligence


Autoria(s): Fernández Martínez, Fernando; Zablotskaya, Kseniya; Minker, Wolfgang
Data(s)

01/08/2012

Resumo

In this paper we investigate whether conventional text categorization methods may suffice to infer different verbal intelligence levels. This research goal relies on the hypothesis that the vocabulary that speakers make use of reflects their verbal intelligence levels. Automatic verbal intelligence estimation of users in a spoken language dialog system may be useful when defining an optimal dialog strategy by improving its adaptation capabilities. The work is based on a corpus containing descriptions (i.e. monologs) of a short film by test persons yielding different educational backgrounds and the verbal intelligence scores of the speakers. First, a one-way analysis of variance was performed to compare the monologs with the film transcription and to demonstrate that there are differences in the vocabulary used by the test persons yielding different verbal intelligence levels. Then, for the classification task, the monologs were represented as feature vectors using the classical TF–IDF weighting scheme. The Naive Bayes, k-nearest neighbors and Rocchio classifiers were tested. In this paper we describe and compare these classification approaches, define the optimal classification parameters and discuss the classification results obtained.

Formato

application/pdf

Identificador

http://oa.upm.es/15937/

Idioma(s)

eng

Publicador

E.T.S.I. Telecomunicación (UPM)

Relação

http://oa.upm.es/15937/1/INVE_MEM_2012_131531.pdf

http://www.sciencedirect.com/science/article/pii/S0957417412004368

info:eu-repo/semantics/altIdentifier/doi/10.1016/j.eswa.2012.02.173

Direitos

http://creativecommons.org/licenses/by-nc-nd/3.0/es/

info:eu-repo/semantics/openAccess

Fonte

Expert Systems with Applications, ISSN 0957-4174, 2012-08, Vol. 39, No. 10

Palavras-Chave #Telecomunicaciones
Tipo

info:eu-repo/semantics/article

Artículo

PeerReviewed