Learning with multiple pairwise kernels for drug bioactivity prediction - UTU Tutkimustietojärjestelmä

Vertaisarvioitu alkuperäisartikkeli tai data-artikkeli tieteellisessä aikakauslehdessä (A1)

Learning with multiple pairwise kernels for drug bioactivity prediction

Julkaisun tekijät: Cichonska A, Pahikkala T, Szedmak S, Julkunen H, Airola A, Heinonen M, Aittokallio T, Rousu J

Kustantaja: OXFORD UNIV PRESS

Julkaisuvuosi: 2018

Journal: Bioinformatics

Tietokannassa oleva lehden nimi: BIOINFORMATICS

Lehden akronyymi: BIOINFORMATICS

Volyymi: 34

Julkaisunumero: 13

Aloitussivu: 509

Lopetussivun numero: 518

Sivujen määrä: 10

ISSN: 1367-4803

eISSN: 1460-2059

DOI: http://dx.doi.org/10.1093/bioinformatics/bty277

Verkko-osoite: https://academic.oup.com/bioinformatics/article/34/13/i509/5045738

Rinnakkaistallenteen osoite: https://research.utu.fi/converis/portal/detail/Publication/32838693

Tiivistelmä

Motivation: Many inference problems in bioinformatics, including drug bioactivity prediction, can be formulated as pairwise learning problems, in which one is interested in making predictions for pairs of objects, e.g. drugs and their targets. Kernel-based approaches have emerged as powerful tools for solving problems of that kind, and especially multiple kernel learning (MKL) offers promising benefits as it enables integrating various types of complex biomedical information sources in the form of kernels, along with learning their importance for the prediction task. However, the immense size of pairwise kernel spaces remains a major bottleneck, making the existing MKL algorithms computationally infeasible even for small number of input pairs.Results: We introduce pairwiseMKL, the first method for time- and memory-efficient learning with multiple pairwise kernels. pairwiseMKL first determines the mixture weights of the input pairwise kernels, and then learns the pairwise prediction function. Both steps are performed efficiently without explicit computation of the massive pairwise matrices, therefore making the method applicable to solving large pairwise learning problems. We demonstrate the performance of pairwiseMKL in two related tasks of quantitative drug bioactivity prediction using up to 167 995 bioactivity measurements and 3120 pairwise kernels: (i) prediction of anticancer efficacy of drug compounds across a large panel of cancer cell lines; and (ii) prediction of target profiles of anticancer compounds across their kinome-wide target spaces. We show that pairwiseMKL provides accurate predictions using sparse solutions in terms of selected kernels, and therefore it automatically identifies also data sources relevant for the prediction problem.

Ladattava julkaisu

This is an electronic reprint of the original article.
This reprint may differ from the original in pagination and typographic detail. Please cite the original version.

bty277.pdf