PASCAL - Pattern Analysis, Statistical Modelling and Computational Learning

Weighted Symbols-based Edit Distance for String-Structured Image Classification
Elisa Fromont, Christophe Ducottet, Elisa Fromont, Anne-Claire Legrand and Marc Sebban
In: ECML/PKDD'10 European Conference on Machine Learning and Principles and Practice of Knowledge Discovery in Databases, September 20th to 24th, 2010, Barcelona, Spain.

Abstract

As an alternative to vector representations, a recent trend in image classification suggests to integrate additional structural infor- mation in the description of images in order to enhance classification accuracy. Rather than being represented in a p-dimensional space, im- ages can typically be encoded in the form of strings, trees or graphs and are usually compared either by computing suited metrics such as the (string or tree)-edit distance, or by testing subgraph isomorphism. In this paper, we propose a new way for representing images in the form of strings whose symbols are weighted according to a TF-IDF-based weight- ing scheme, inspired from information retrieval. To be able to handle such real-valued weights, we first introduce a new weighted string edit distance that keeps the properties of a distance. In particular, we prove that the triangle inequality is preserved which allows the computation of the edit distance in quadratic time by dynamic programming. We show on an image classification task that our new weighted edit distance not only significantly outperforms the standard edit distance but also seems very competitive in comparison with standard histogram distances-based approaches.

EPrint Type:Conference or Workshop Item (Paper)
Project Keyword:Project Keyword UNSPECIFIED
Subjects:Machine Vision
ID Code:7403
Deposited By:Elisa Fromont
Deposited On:17 March 2011