Genome-wide association analysis by lasso penalized logistic regression

doi:10.1093/BIOINFORMATICS/BTP041

Open AccessJournal ArticleDOI

Genome-wide association analysis by lasso penalized logistic regression

Tong Tong Wu, +4 more

- 01 Mar 2009 -

Bioinformatics

- Vol. 25, Iss: 6, pp 714-721

TLDR

The performance of lasso penalized logistic regression in case-control disease gene mapping with a large number of SNPs (single nucleotide polymorphisms) predictors is evaluated and coeliac disease results replicate the previous SNP results and shed light on possible interactions among the SNPs.

Abstract:

Motivation: In ordinary regression, imposition of a lasso penalty makes continuous model selection straightforward. Lasso penalized regression is particularly advantageous when the number of predictors far exceeds the number of observations. Method: The present article evaluates the performance of lasso penalized logistic regression in case–control disease gene mapping with a large number of SNPs (single nucleotide polymorphisms) predictors. The strength of the lasso penalty can be tuned to select a predetermined number of the most relevant SNPs and other predictors. For a given value of the tuning constant, the penalized likelihood is quickly maximized by cyclic coordinate ascent. Once the most potent marginal predictors are identified, their two-way and higher order interactions can also be examined by lasso penalized logistic regression. Results: This strategy is tested on both simulated and real data. Our findings on coeliac disease replicate the previous SNP results and shed light on possible interactions among the SNPs. Availability: The software discussed is available in Mendel 9.0 at the UCLA Human Genetics web site. Contact: klange@ucla.edu Supplementary information: Supplementary data are available at Bioinformatics online.

Citations

PDF

Open Access

More filters

Journal ArticleDOI

A novel variational Bayes multiple locus Z-statistic for genome-wide association studies with Bayesian model averaging

Benjamin A. Logsdon, +4 more

- 01 Jul 2012 -

Bioinformatics

TL;DR: This methodology is the first penalized multiple regression approach that explicitly controls Type I error rates and provide model over-fitting diagnostics through a novel normally distributed statistic defined for every marker within the GWAS, based on results from a variational Bayes spike regression algorithm.

...read moreread less

Journal ArticleDOI

Minimum Distance Lasso for robust high-dimensional regression

Aurelie C. Lozano, +2 more

- 01 Jan 2016 -

Electronic Journal of Statistics

TL;DR: The method, Minimum Distance Lasso (MD-Lasso), combines minimum distance functionals customarily used in nonparametric estimation for robustness, with 1-regularization, and establishes a connection with re-weighted least-squares that intuitively explains MD- Lasso robustness.

...read moreread less

Journal ArticleDOI

Bag of Naïve Bayes: biomarker selection and classification from genome-wide SNP data

Francesco Sambo, +4 more

- 07 Sep 2012 -

BMC Bioinformatics

TL;DR: The significantly higher classification accuracy obtained by BoNB, together with the significance of the biomarkers identified from the Type 1 Diabetes dataset, prove the effectiveness of BoNB as an algorithm for both classification and biomarker selection from genome-wide SNP data.

...read moreread less

Journal ArticleDOI

Feature ranking of active region source properties in solar flare forecasting and the uncompromised stochasticity of flare occurrence

Cristina Campi, +5 more

- 28 Jun 2019 -

arXiv: Solar and Stellar Astrophysics

TL;DR: In this paper, the authors utilized an unprecedented 171 flare-predictive active region properties, mainly inferred by the Helioseismic and Magnetic Imager onboard the Solar Dynamics Observatory (SDO/HMI) in the course of the European Union Horizon 2020 FLARECAST project.

...read moreread less

Journal ArticleDOI

Multiple-kernel learning for genomic data mining and prediction

Chris Wilson, +4 more

- 15 Aug 2019 -

BMC Bioinformatics

TL;DR: It is shown that MKL can identify gene sets that are known to play a role in the prognostic prediction of 15 cancer types using gene expression data from The Cancer Genome Atlas, as well as, identify new gene sets for the future research.

...read moreread less

Collapse

References

PDF

Open Access

More filters

Journal ArticleDOI

Controlling the false discovery rate: a practical and powerful approach to multiple testing

Yoav Benjamini, +1 more

- 01 Jan 1995 -

Journal of the royal statistical society...

TL;DR: In this paper, a different approach to problems of multiple significance testing is presented, which calls for controlling the expected proportion of falsely rejected hypotheses -the false discovery rate, which is equivalent to the FWER when all hypotheses are true but is smaller otherwise.

...read moreread less

Journal ArticleDOI

Regression Shrinkage and Selection via the Lasso

Robert Tibshirani

- 01 Jan 1996 -

Journal of the royal statistical society...

TL;DR: A new method for estimation in linear models called the lasso, which minimizes the residual sum of squares subject to the sum of the absolute value of the coefficients being less than a constant, is proposed.

...read moreread less

Journal ArticleDOI

Regularization Paths for Generalized Linear Models via Coordinate Descent

Jerome H. Friedman, +2 more

- 02 Feb 2010 -

Journal of Statistical Software

TL;DR: In comparative timings, the new algorithms are considerably faster than competing methods and can handle large problems and can also deal efficiently with sparse features.

...read moreread less

Journal ArticleDOI

Atomic Decomposition by Basis Pursuit

Scott Chen, +2 more

- 11 Dec 1998 -

SIAM Journal on Scientific Computing

TL;DR: Basis Pursuit (BP) is a principle for decomposing a signal into an "optimal" superposition of dictionary elements, where optimal means having the smallest l1 norm of coefficients among all such decompositions.

...read moreread less

Journal ArticleDOI

An Iterative Thresholding Algorithm for Linear Inverse Problems with a Sparsity Constraint

Ingrid Daubechies, +2 more

- 01 Nov 2004 -

Communications on Pure and Applied Mathe...

TL;DR: It is proved that replacing the usual quadratic regularizing penalties by weighted 𝓁p‐penalized penalties on the coefficients of such expansions, with 1 ≤ p ≤ 2, still regularizes the problem.

...read moreread less