Finding regulatory modules through large-scale gene-expression data analysis
Morten Kloster, Chao Tang, Ned Wingreen
Abstract
The use of gene microchips has enabled a rapid accumulation of gene-expression data. One of the major challenges of analyzing this data is the diversity, in both size and signal strength, of the various modules in the gene regulatory networks of organisms. Based on the Iterative Signature Algorithm [Bergmann, S., Ihmels, J. and Barkai, N. (2002) Phys. Rev. E 67, 031902], we present an algorithm - the Progressive Iterative Signature Algorithm (PISA) - that, by sequentially eliminating modules, allows unsupervised identification of both large and small regulatory modules. We applied PISA to a large set of yeast gene-expression data, and, using the Gene Ontology annotation database as a reference, found that our algorithm is much better able to identify regulatory modules than methods based on high-throughput transcription-factor binding experiments or on comparative genomics.
Create a lesson
Related papers
Automatic denoising and differentiation based on Savitzky-Golay filtering and Homogeneous Differentiators for attractor reconstruction via differential embedding
Uros Sutulovic, Daniele Proverbio, Rami Katz et al.
FlowLOT: Linearized Optimal Transport for Flow Cytometry Analysis
Naqib Sad Pathan, Mohammad Shifat-E-Rabbi, Kristofor E. Pas et al.
GIA: Germline-Informed Aging with AlphaGenome Finds Genetically Regulated CpGs
Sean Lim
Decoding Extrahepatic Targeting of Lipid Nanoparticles with Interpretable Machine Learning
Asal Mehradfar, Mohammad Shahab Sepehri, Owen Antholine et al.
GPCR Ligand Bioactivity Prediction with Physics-Informed Dual-State Query Learning
Shuo Zhang, Huifeng Zhang, Rongqi Hong et al.
Optical microelectrode arrays for differential readout of electrical and mechanical signals in cardiac cells
Alessandro Leronni, Rosalia Moreddu