The Community for Technology Leaders
RSS Icon
Subscribe
Lyon
Aug. 22, 2011 to Aug. 27, 2011
ISBN: 978-1-4577-1373-6
pp: 217-220
ABSTRACT
This paper proposes a two-phase feature selection method specific for bioinformatics domain from classification perspective in data mining. In the first phase, Bhattacharyya distance measurement is used for filtering the majority of irrelevant genes. Upon the basis, we apply floating sequential search method (FSSM) to further select informative gene set using kernel distance as measurement of class separability. The verification of colon tissue dataset using support vector machines (SVMs) proves that informative gene set selected by our method is acceptable for disease identification.
INDEX TERMS
feature selection, domain driven data mining, kernel distance measurement, Bhattacharyya distance, floating sequential search method
CITATION
Lingling Zhang, Jun Li, Yibing Chen, "Domain Driven Two-Phase Feature Selection Method Based on Bhattacharyya Distance and Kernel Distance Measurements", WI-IAT, 2011, 2011 IEEE/WIC/ACM International Joint Conferences on Web Intelligence (WI) and Intelligent Agent Technologies, 2011 IEEE/WIC/ACM International Joint Conferences on Web Intelligence (WI) and Intelligent Agent Technologies 2011, pp. 217-220, doi:10.1109/WI-IAT.2011.61
10 ms
(Ver 2.0)

Marketing Automation Platform Marketing Automation Tool