Abstract
Recent tools that analyze microarray expression data have exploited correlation-based approaches such as clustering analysis. We describe a new method for assessing the importance of genes for sample classification based on expression data. Our approach combines a genetic algorithm (GA) and the k-nearest neighbor (KNN) method to identify genes that jointly can discriminate between two types of samples (e.g. normal vs. tumor). First, many such subsets of differentially expressed genes are obtained independently using the GA. Then, the overall frequency with which genes were selected is used to deduce the relative importance of genes for sample classification. Sample heterogeneity is accommodated; that is, the method should be robust against the existence of distinct subtypes. We applied GA / KNN to expression data from normal versus tumor tissue from human colon. Two distinct clusters were observed when the 50 most frequently selected genes were used to classify all of the samples in the data sets stu died and the majority of samples were classified correctly. Identification of a set of differentially expressed genes could aid in tumor diagnosis and could also serve to identify disease subtypes that may benefit from distinct clinical approaches to treatment.
Keywords: Gene Expression, Algorithm (GA), K-nearest neighbor (KNN), Pattern recognition, Gene selection, High-dimensional, Microarray
Combinatorial Chemistry & High Throughput Screening
Title: Gene Assessment and Sample Classification for Gene Expression Data Using a Genetic Algorithm / k-nearest Neighbor Method
Volume: 4 Issue: 8
Author(s): Leping Li, Thomas A. Darden, Clarice R. Weingberg, A. J. Levine and Lee G. Pedersen
Affiliation:
Keywords: Gene Expression, Algorithm (GA), K-nearest neighbor (KNN), Pattern recognition, Gene selection, High-dimensional, Microarray
Abstract: Recent tools that analyze microarray expression data have exploited correlation-based approaches such as clustering analysis. We describe a new method for assessing the importance of genes for sample classification based on expression data. Our approach combines a genetic algorithm (GA) and the k-nearest neighbor (KNN) method to identify genes that jointly can discriminate between two types of samples (e.g. normal vs. tumor). First, many such subsets of differentially expressed genes are obtained independently using the GA. Then, the overall frequency with which genes were selected is used to deduce the relative importance of genes for sample classification. Sample heterogeneity is accommodated; that is, the method should be robust against the existence of distinct subtypes. We applied GA / KNN to expression data from normal versus tumor tissue from human colon. Two distinct clusters were observed when the 50 most frequently selected genes were used to classify all of the samples in the data sets stu died and the majority of samples were classified correctly. Identification of a set of differentially expressed genes could aid in tumor diagnosis and could also serve to identify disease subtypes that may benefit from distinct clinical approaches to treatment.
Export Options
About this article
Cite this article as:
Li Leping, Darden A. Thomas, Weingberg R. Clarice, Levine J. A. and Pedersen G. Lee, Gene Assessment and Sample Classification for Gene Expression Data Using a Genetic Algorithm / k-nearest Neighbor Method, Combinatorial Chemistry & High Throughput Screening 2001; 4 (8) . https://dx.doi.org/10.2174/1386207013330733
DOI https://dx.doi.org/10.2174/1386207013330733 |
Print ISSN 1386-2073 |
Publisher Name Bentham Science Publisher |
Online ISSN 1875-5402 |
Call for Papers in Thematic Issues
Artificial Intelligence Methods for Biomedical, Biochemical and Bioinformatics Problems
Recently, a large number of technologies based on artificial intelligence have been developed and applied to solve a diverse range of problems in the areas of biomedical, biochemical and bioinformatics problems. By utilizing powerful computing resources and massive amounts of data, methods based on artificial intelligence can significantly improve the ...read more
Eco-friendly Agents for Biological Control of Pathogenic Diseases
The discovery of an alternative biological approach to disease management includes work on medicinal products derived from natural sources as a starting point for the development of eco-friendly agents for these diseases and the injuries they cause, as well as reducing human contact with hazardous chemicals and their residues. We ...read more
Emerging trends in diseases mechanisms, noble drug targets and therapeutic strategies: focus on immunological and inflammatory disorders
Recently infectious and inflammatory diseases have been a key concern worldwide due to tremendous morbidity and mortality world Wide. Recent, nCOVID-9 pandemic is a good example for the emerging infectious disease outbreak. The world is facing many emerging and re-emerging diseases out breaks at present however, there is huge lack ...read more
Exploring Spectral Graph Theory in Combinatorial Chemistry
Scope of the Thematic Issue: Combinatorial chemistry involves the synthesis and analysis of a large number of diverse compounds simultaneously. Traditional methods rely on brute force experimentation, which can be time-consuming and resource-intensive. Spectral Graph Theory, a branch of mathematics dealing with the properties of graphs in relation to the ...read more
- Author Guidelines
- Graphical Abstracts
- Fabricating and Stating False Information
- Research Misconduct
- Post Publication Discussions and Corrections
- Publishing Ethics and Rectitude
- Increase Visibility of Your Article
- Archiving Policies
- Peer Review Workflow
- Order Your Article Before Print
- Promote Your Article
- Manuscript Transfer Facility
- Editorial Policies
- Allegations from Whistleblowers
Related Articles
-
Prodrugs for Targeted Tumor Therapies: Recent Developments in ADEPT, GDEPT and PMT
Current Pharmaceutical Design B Lymphocyte Memory in X-Linked Lymphoproliferative Disease (XLP)
Current Immunology Reviews (Discontinued) Herpesvirus Saimiri-Based Gene Delivery Vectors
Current Gene Therapy Occurrence and Biological Properties of Sphingolipids - A Review
Current Nutrition & Food Science The Synthetic and Biological Attributes of Pyrazole Derivatives: A Review
Mini-Reviews in Medicinal Chemistry Development of Novel Therapeutics Targeting Isocitrate Dehydrogenase Mutations in Cancer
Current Topics in Medicinal Chemistry Anti - TNF Therapy for Crohns Disease
Current Pharmaceutical Design Folate-conjugated Chitosan-poly(ethylenimine) Copolymer As An Efficient and Safe Vector For Gene Delivery in Cancer Cells
Current Gene Therapy Therapy of Elderly/Comorbid Patients with Chronic Lymphocytic Leukemia
Current Pharmaceutical Design Therapeutic Monoclonal Antibodies: Strategies and Challenges for Biosimilars Development
Current Medicinal Chemistry Histone Deacetylase Inhibitors: Molecular and Biological Activity as a Premise to Clinical Application
Current Drug Metabolism Recent Advances in Carbon Nanotubes as Delivery Systems for Anticancer Drugs
Current Medicinal Chemistry Discovery of Small Molecule c-Met Inhibitors: Evolution and Profiles of Clinical Candidates
Anti-Cancer Agents in Medicinal Chemistry Histone Lysine-Specific Methyltransferases and Demethylases in Carcinogenesis: New Targets for Cancer Therapy and Prevention
Current Cancer Drug Targets Morpho-Functional Features of In-Vitro Cell Death Induced by Physical Agents
Current Pharmaceutical Design Current Developments in the Therapeutic Potential of S-Nitrosoglutathione, an Endogenous NO-Donor Molecule
Current Pharmaceutical Biotechnology Neurodegenerative Disease: A Perspective on Cell-Based Therapy in the New Era of Cell-Free Nano-Therapy
Current Pharmaceutical Design Novel Chemotherapeutic Agents - The Contribution of Scorpionates
Current Medicinal Chemistry Testicular Germ Cell Tumors: A Paradigm for the Successful Treatment of Solid Tumor Stem Cells
Current Cancer Therapy Reviews Synthetic 2-Methoxyestradiol Derivatives: Structure-Activity Relationships
Current Medicinal Chemistry