Learning with multiple pairwise kernels for drug bioactivity prediction - UTU Research Portal

A1 Refereed original research article in a scientific journal

Learning with multiple pairwise kernels for drug bioactivity prediction

Authors: Cichonska A, Pahikkala T, Szedmak S, Julkunen H, Airola A, Heinonen M, Aittokallio T, Rousu J

Publisher: OXFORD UNIV PRESS

Publication year: 2018

Journal: Bioinformatics

Journal name in source: BIOINFORMATICS

Journal acronym: BIOINFORMATICS

Volume: 34

Issue: 13

First page : 509

Last page: 518

Number of pages: 10

ISSN: 1367-4803

eISSN: 1460-2059

DOI: https://doi.org/10.1093/bioinformatics/bty277

Web address : https://academic.oup.com/bioinformatics/article/34/13/i509/5045738

Self-archived copy’s web address: https://research.utu.fi/converis/portal/detail/Publication/32838693

Abstract

Motivation: Many inference problems in bioinformatics, including drug bioactivity prediction, can be formulated as pairwise learning problems, in which one is interested in making predictions for pairs of objects, e.g. drugs and their targets. Kernel-based approaches have emerged as powerful tools for solving problems of that kind, and especially multiple kernel learning (MKL) offers promising benefits as it enables integrating various types of complex biomedical information sources in the form of kernels, along with learning their importance for the prediction task. However, the immense size of pairwise kernel spaces remains a major bottleneck, making the existing MKL algorithms computationally infeasible even for small number of input pairs.Results: We introduce pairwiseMKL, the first method for time- and memory-efficient learning with multiple pairwise kernels. pairwiseMKL first determines the mixture weights of the input pairwise kernels, and then learns the pairwise prediction function. Both steps are performed efficiently without explicit computation of the massive pairwise matrices, therefore making the method applicable to solving large pairwise learning problems. We demonstrate the performance of pairwiseMKL in two related tasks of quantitative drug bioactivity prediction using up to 167 995 bioactivity measurements and 3120 pairwise kernels: (i) prediction of anticancer efficacy of drug compounds across a large panel of cancer cell lines; and (ii) prediction of target profiles of anticancer compounds across their kinome-wide target spaces. We show that pairwiseMKL provides accurate predictions using sparse solutions in terms of selected kernels, and therefore it automatically identifies also data sources relevant for the prediction problem.

Downloadable publication

This is an electronic reprint of the original article.
This reprint may differ from the original in pagination and typographic detail. Please cite the original version.

bty277.pdf