scispace - formally typeset
Open AccessJournal ArticleDOI

Clustering Algorithms: Their Application to Gene Expression Data

Reads0
Chats0
TLDR
This review examines the various clustering algorithms applicable to the gene expression data in order to discover and provide useful knowledge of the appropriate clustering technique that will guarantee stability and high degree of accuracy in its analysis procedure.
Abstract
Gene expression data hide vital information required to understand the biological process that takes place in a particular organism in relation to its environment. Deciphering the hidden patterns in gene expression data proffers a prodigious preference to strengthen the understanding of functional genomics. The complexity of biological networks and the volume of genes present increase the challenges of comprehending and interpretation of the resulting mass of data, which consists of millions of measurements; these data also inhibit vagueness, imprecision, and noise. Therefore, the use of clustering techniques is a first step toward addressing these challenges, which is essential in the data mining process to reveal natural structures and identify interesting patterns in the underlying data. The clustering of gene expression data has been proven to be useful in making known the natural structure inherent in gene expression data, understanding gene functions, cellular processes, and subtypes of cells, mining useful information from noisy data, and understanding gene regulation. The other benefit of clustering gene expression data is the identification of homology, which is very important in vaccine design. This review examines the various clustering algorithms applicable to the gene expression data in order to discover and provide useful knowledge of the appropriate clustering technique that will guarantee stability and high degree of accuracy in its analysis procedure.

read more

Citations
More filters

The Self-Organizing Map

TL;DR: An overview of the self-organizing map algorithm, on which the papers in this issue are based, is presented in this article, where the authors present an overview of their work.
Journal ArticleDOI

Applications of machine learning to diagnosis and treatment of neurodegenerative diseases

TL;DR: How machine learning can aid early diagnosis and interpretation of medical images as well as the discovery and development of new therapies is discussed, and the latest developments in the use of machine learning to interrogate neurodegenerative disease-related datasets are described.
Journal ArticleDOI

A comprehensive survey of clustering algorithms: State-of-the-art machine learning applications, taxonomy, challenges, and future research prospects

TL;DR: Clustering is an essential tool in data mining research and applications as discussed by the authors and it is the subject of active research in many fields of study, such as computer science, data science, statistics, pattern recognition, artificial intelligence, and machine learning.
Journal ArticleDOI

Deep learning-based clustering approaches for bioinformatics

TL;DR: In this article, the authors present a review of state-of-the-art DL-based approaches for clustering analysis that are based on representation learning, which they hope to be useful for bioinformatics research.
Journal ArticleDOI

Single-cell transcriptomic evidence for dense intracortical neuropeptide networks

TL;DR: Here, neuron-type-specific patterns of NP gene expression are used to offer specific, testable predictions regarding 37 peptidergic neuromodulatory networks that may play prominent roles in cortical homeostasis and plasticity.
References
More filters
Journal ArticleDOI

A quantitative study of gene regulation involved in the immune response of Anopheline mosquitoes: An application of Bayesian hierarchical clustering of curves

TL;DR: A Bayesian model-based hierarchical clustering algorithm is introduced for curve data to investigate mechanisms of regulation in the genes concerned and reveals structure within the data not captured by other approaches.
Journal ArticleDOI

A cluster validity index for fuzzy clustering

TL;DR: A new validity index is proposed that employs a compactness measure and a separation measure and shows the superior effectiveness and reliability of the proposed index in comparison to other indices.
Journal ArticleDOI

An improved algorithm for clustering gene expression data

TL;DR: The significant superiority of the proposed two-stage clustering algorithm as compared to the average linkage method, Self Organizing Map (SOM) and a recently developed weighted Chinese restaurant-based clustering method (CRC), widely used methods for clustering gene expression data, is established.
Journal ArticleDOI

An automatic method to determine the number of clusters using decision-theoretic rough set

TL;DR: An efficient automatic method by extending the decision-theoretic rough set model to clustering, which is proved to stop automatically at the perfect number of clusters without manual interference, and a novel fast algorithm, FACA-DTRS, is devised based on the conclusion obtained in the validation of the ACA-D TRS algorithm.
Journal ArticleDOI

Genomic DNA standards for gene expression profiling in Mycobacterium tuberculosis

TL;DR: Testing a normalization procedure based on comparing gene expression levels to the signals generated from hybridizing genomic DNA concluded that genomic DNA standards offer advantages over conventional RNA normalization procedures and can be adapted for the investigation of microbial genomes.
Related Papers (5)
Trending Questions (1)
What are applications of clustering algorithms?

Applications of clustering algorithms include revealing natural structures in gene expression data, understanding gene functions, identifying cell subtypes, mining information from noisy data, and aiding in vaccine design.