Please use this identifier to cite or link to this item: http://repositorio.inesctec.pt/handle/123456789/2829
Title: D-Confidence: an active learning strategy to reduce label disclosure complexity in the presence of imbalanced class distributions
Authors: Alípio Jorge
Nuno Escudeiro
Issue Date: 2012
Abstract: In some classification tasks, such as those related to the automatic building and maintenance of text corpora, it is expensive to obtain labeled instances to train a clas- sifier. In such circumstances it is common to have mas- sive corpora where a few instances are labeled (typically a minority) while others are not. Semi-supervised learning techniques try to leverage the intrinsic information in unla- beled instances to improve classification models. However, these techniques assume that the labeled instances cover all the classes to learn which might not be the case. More- over, when in the presence of an imbalanced class distribution, getting labeled instances from minority classes might be very costly, requiring extensive labeling, if queries are randomly selected. Active learning allows asking an oracle to label new instances, which are selected by criteria, aiming to reduce the labeling effort. D-Confidence is an active learning approach that is effective when in pres- enc
URI: http://repositorio.inesctec.pt/handle/123456789/2829
http://dx.doi.org/10.1007/s13173-012-0069-3
metadata.dc.type: article
Publication
Appears in Collections:LIAAD - Articles in International Journals

Files in This Item:
File Description SizeFormat 
PS-08366.pdf1.28 MBAdobe PDFView/Open


Items in DSpace are protected by copyright, with all rights reserved, unless otherwise indicated.