Del Favero, Simone - Varagnolo, Damiano - Dinuzzo, Francesco - Schenato, Luca - Pillonetto, Gianluigi (2011) On the discardability of data in Support Vector Classification problems. In: IEEE Conference on Decision and Control, 12 - 15 December, 2011, Orlando, Florida, USA.

Abstract (inglese)

We analyze the problem of data sets reduction for support vector classification. The work is also motivated by distributed problems, where sensors collect binary measurements at different locations moving inside an environment that needs to be divided into a collection of regions labeled in two different ways. The scope is to let each agent retain and exchange only those measurements that are mostly informative for the collective reconstruction of the decision boundary. For the case of separable classes, we provide the exact conditions and an efficient algorithm to determine if an element in the training set can become a support vector when new data arrive. The analysis is then extended to the non-separable case deriving a sufficient discardability condition and a general data selection scheme for classification. Numerical experiments relative to the distributed problem show that the proposed procedure allows the agents to exchange a small amount of the collected data to obtain a highly predictive decision boundary.

Tipo di EPrint:Contributo a convegno (Relazione)
Anno di Pubblicazione:12 Dicembre 2011
Parole chiave (italiano / inglese):distributed classification, support vector machines, model reduction, distributed machine learning, simplex, convex analysis
Settori scientifico-disciplinari MIUR:Area 09 - Ingegneria industriale e dell'informazione > ING-INF/04 Automatica
Struttura di riferimento:Dipartimenti > Dipartimento di Ingegneria dell'Informazione
Codice ID:4302
Depositato il:29 Set 2011 15:18
