Abstract
Recent tools that analyze microarray expression data have exploited correlation-based approaches such as clustering analysis. We describe a new method for assessing the importance of genes for sample classification based on expression data. Our approach combines a genetic algorithm (GA) and the k-nearest neighbor (KNN) method to identify genes that jointly can discriminate between two types of samples (e.g. normal vs. tumor). First, many such subsets of differentially expressed genes are obtained independently using the GA. Then, the overall frequency with which genes were selected is used to deduce the relative importance of genes for sample classification. Sample heterogeneity is accommodated; that is, the method should be robust against the existence of distinct subtypes. We applied GA / KNN to expression data from normal versus tumor tissue from human colon. Two distinct clusters were observed when the 50 most frequently selected genes were used to classify all of the samples in the data sets stu died and the majority of samples were classified correctly. Identification of a set of differentially expressed genes could aid in tumor diagnosis and could also serve to identify disease subtypes that may benefit from distinct clinical approaches to treatment.
Keywords: Gene Expression, Algorithm (GA), K-nearest neighbor (KNN), Pattern recognition, Gene selection, High-dimensional, Microarray
Combinatorial Chemistry & High Throughput Screening
Title: Gene Assessment and Sample Classification for Gene Expression Data Using a Genetic Algorithm / k-nearest Neighbor Method
Volume: 4 Issue: 8
Author(s): Leping Li, Thomas A. Darden, Clarice R. Weingberg, A. J. Levine and Lee G. Pedersen
Affiliation:
Keywords: Gene Expression, Algorithm (GA), K-nearest neighbor (KNN), Pattern recognition, Gene selection, High-dimensional, Microarray
Abstract: Recent tools that analyze microarray expression data have exploited correlation-based approaches such as clustering analysis. We describe a new method for assessing the importance of genes for sample classification based on expression data. Our approach combines a genetic algorithm (GA) and the k-nearest neighbor (KNN) method to identify genes that jointly can discriminate between two types of samples (e.g. normal vs. tumor). First, many such subsets of differentially expressed genes are obtained independently using the GA. Then, the overall frequency with which genes were selected is used to deduce the relative importance of genes for sample classification. Sample heterogeneity is accommodated; that is, the method should be robust against the existence of distinct subtypes. We applied GA / KNN to expression data from normal versus tumor tissue from human colon. Two distinct clusters were observed when the 50 most frequently selected genes were used to classify all of the samples in the data sets stu died and the majority of samples were classified correctly. Identification of a set of differentially expressed genes could aid in tumor diagnosis and could also serve to identify disease subtypes that may benefit from distinct clinical approaches to treatment.
Export Options
About this article
Cite this article as:
Li Leping, Darden A. Thomas, Weingberg R. Clarice, Levine J. A. and Pedersen G. Lee, Gene Assessment and Sample Classification for Gene Expression Data Using a Genetic Algorithm / k-nearest Neighbor Method, Combinatorial Chemistry & High Throughput Screening 2001; 4 (8) . https://dx.doi.org/10.2174/1386207013330733
DOI https://dx.doi.org/10.2174/1386207013330733 |
Print ISSN 1386-2073 |
Publisher Name Bentham Science Publisher |
Online ISSN 1875-5402 |
Call for Papers in Thematic Issues
Advances in the design of antibody & protein with conformational dynamics and artificial intelligence approaches
“Antibodies & Protein Design” section focuses on the utilization of multiple strategies to engineer and optimize antibodies and proteins that serve diverse analytical strategies, such as combinatorial protein design, structure-based design, sequence-based design, and other techniques that incorporate principles of protein-protein interactions, allosteric regulation, and post-translational modifications. Example applications include ...read more
Artificial Intelligence Methods for Biomedical, Biochemical and Bioinformatics Problems
Recently, a large number of technologies based on artificial intelligence have been developed and applied to solve a diverse range of problems in the areas of biomedical, biochemical and bioinformatics problems. By utilizing powerful computing resources and massive amounts of data, methods based on artificial intelligence can significantly improve the ...read more
Emerging trends in diseases mechanisms, noble drug targets and therapeutic strategies: focus on immunological and inflammatory disorders
Recently infectious and inflammatory diseases have been a key concern worldwide due to tremendous morbidity and mortality world Wide. Recent, nCOVID-9 pandemic is a good example for the emerging infectious disease outbreak. The world is facing many emerging and re-emerging diseases out breaks at present however, there is huge lack ...read more
Exploring Spectral Graph Theory in Combinatorial Chemistry
Combinatorial chemistry involves the synthesis and analysis of a large number of diverse compounds simultaneously. Traditional methods rely on brute-force experimentation, which can be time-consuming and resource-intensive. Spectral graph theory, a branch of mathematics dealing with the properties of graphs in relation to the eigenvalues and eigenvectors of matrices associated ...read more
- Author Guidelines
- Graphical Abstracts
- Fabricating and Stating False Information
- Research Misconduct
- Post Publication Discussions and Corrections
- Publishing Ethics and Rectitude
- Increase Visibility of Your Article
- Archiving Policies
- Peer Review Workflow
- Order Your Article Before Print
- Promote Your Article
- Manuscript Transfer Facility
- Editorial Policies
- Allegations from Whistleblowers
Related Articles
-
Natural Flora and Anticancer Regime: Milestones and Roadmap
Anti-Cancer Agents in Medicinal Chemistry Epigenetic Therapies of Cancer
Current Cancer Therapy Reviews CCL21 and IFNγ Recruit and Activate Tumor Specific T cells in 3D Scaffold Model of Breast Cancer
Anti-Cancer Agents in Medicinal Chemistry DNA Methyltransferase Inhibitors and their Therapeutic Potential
Current Topics in Medicinal Chemistry Safer Vectors for Gene Therapy of Primary Immunodeficiencies
Current Gene Therapy Antiangiogenic Therapies in Non-Hodgkin's Lymphoma
Current Cancer Drug Targets Pharmacological Modulation of Caspase Activation
Current Medicinal Chemistry - Anti-Inflammatory & Anti-Allergy Agents Arylpyrazoles: Heterocyclic Scaffold of Immense Therapeutic Application
Current Organic Chemistry Understanding Molecular Process and Chemotherapeutics for the Management of Breast Cancer
Current Chemical Biology Targeting the PI3K/AKT/mTOR Signaling Pathway in Primary Central Nervous System Lymphoma: Current Status and Future Prospects
CNS & Neurological Disorders - Drug Targets Targeting Bcl-2 in CLL
Current Medicinal Chemistry Role of Imaging in Testicular Cancer
Current Medical Imaging Platinum-Intercalator Conjugates: From DNA-Targeted Cisplatin Derivatives to Adenine Binding Complexes as Potential Modulators of Gene Regulation
Current Topics in Medicinal Chemistry Anti-Cancer Therapeutic Approaches Based on Intracellular and Extracellular Heat Shock Proteins
Current Medicinal Chemistry Clinical Implications of Methotrexate Pharmacogenetics in Childhood Acute Lymphoblastic Leukaemia
Current Drug Metabolism Anticancer Potential of Biologically Active Diosgenin and its Derivatives: An Update
Current Traditional Medicine An Overview on 2-arylquinolin-4(1H)-ones and Related Structures as Tubulin Polymerisation Inhibitors
Current Topics in Medicinal Chemistry Recommendations for Severe Hypertriglyceridemia Treatment, are there New Strategies?
Current Vascular Pharmacology Ceramide and Apoptosis: Exploring the Enigmatic Connections between Sphingolipid Metabolism and Programmed Cell Death
Anti-Cancer Agents in Medicinal Chemistry Advances in Cancer Stem Cell Therapy: Targets and Treatments
Recent Patents on Regenerative Medicine