Data visualization with simultaneous feature selection

Research output: Chapter in Book/Report/Conference proceedingConference contribution

View graph of relations Save citation

Open

Authors

Research units

Abstract

Data visualization algorithms and feature selection techniques are both widely used in bioinformatics but as distinct analytical approaches. Until now there has been no method of measuring feature saliency while training a data visualization model. We derive a generative topographic mapping (GTM) based data visualization approach which estimates feature saliency simultaneously with the training of the visualization model. The approach not only provides a better projection by modeling irrelevant features with a separate noise model but also gives feature saliency values which help the user to assess the significance of each feature. We compare the quality of projection obtained using the new approach with the projections from traditional GTM and self-organizing maps (SOM) algorithms. The results obtained on a synthetic and a real-life chemoinformatics dataset demonstrate that the proposed approach successfully identifies feature significance and provides coherent (compact) projections. © 2006 IEEE.

Documents

  • NCRG_2006_014.pdf

    Rights statement: © 2006 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works.

    Accepted author manuscript, 227 KB, PDF-document

Details

Publication date2006
Publication titleProceedings of the 2006 IEEE Symposium on Computational Intelligence in Bioinformatics and Computational Biology, CIBCB'06
Pages156-163
Number of pages8
Original languageEnglish
Event3rd symposium on Computational Intelligence in Bioinformatics and Computational Biology - Toronto, ON, Canada

Symposium

Symposium3rd symposium on Computational Intelligence in Bioinformatics and Computational Biology
Abbreviated titleCIBCB '06
CountryCanada
CityToronto, ON
Period28/09/0629/09/06

Bibliographic note

© 2006 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works.

    Keywords

  • chemoinformatics, data mining, data visualization, feature selection, generative topographic mapping, unsupervised learning

DOI

Download statistics

No data available

Employable Graduates; Exploitable Research

Copy the text from this field...