QSAR-derived affinity fingerprints (part 2): modeling performance for potency prediction

Identifikátory výsledku

Kód výsledku v IS VaVaI
<a href="https://www.isvavai.cz/riv?ss=detail&h=RIV%2F68378050%3A_____%2F20%3A00538137" target="_blank" >RIV/68378050:_____/20:00538137 - isvavai.cz</a>
Nalezeny alternativní kódy
RIV/60461373:22310/20:43921545
Výsledek na webu
<a href="https://jcheminf.biomedcentral.com/articles/10.1186/s13321-020-00444-5" target="_blank" >https://jcheminf.biomedcentral.com/articles/10.1186/s13321-020-00444-5</a>
DOI - Digital Object Identifier
<a href="http://dx.doi.org/10.1186/s13321-020-00444-5" target="_blank" >10.1186/s13321-020-00444-5</a>

Alternativní jazyky

Jazyk výsledku
angličtina
Název v původním jazyce
QSAR-derived affinity fingerprints (part 2): modeling performance for potency prediction
Popis výsledku v původním jazyce
Affinity fingerprints report the activity of small molecules across a set of assays, and thus permit to gather information about the bioactivities of structurally dissimilar compounds, where models based on chemical structure alone are often limited, and model complex biological endpoints, such as human toxicity and in vitro cancer cell line sensitivity. Here, we propose to model in vitro compound activity using computationally predicted bioactivity profiles as compound descriptors. To this aim, we apply and validate a framework for the calculation of QSAR-derived affinity fingerprints (QAFFP) using a set of 1360 QSAR models generated using K-i, K-d, IC50 and EC50 data from ChEMBL database. QAFFP thus represent a method to encode and relate compounds on the basis of their similarity in bioactivity space. To benchmark the predictive power of QAFFP we assembled IC50 data from ChEMBL database for 18 diverse cancer cell lines widely used in preclinical drug discovery, and 25 diverse protein target data sets. This study complements part 1 where the performance of QAFFP in similarity searching, scaffold hopping, and bioactivity classification is evaluated. Despite being inherently noisy, we show that using QAFFP as descriptors leads to errors in prediction on the test set in the similar to 0.65-0.95 pIC(50) units range, which are comparable to the estimated uncertainty of bioactivity data in ChEMBL (0.76-1.00 pIC(50) units). We find that the predictive power of QAFFP is slightly worse than that of Morgan2 fingerprints and 1D and 2D physicochemical descriptors, with an effect size in the 0.02-0.08 pIC(50) units range. Including QSAR models with low predictive power in the generation of QAFFP does not lead to improved predictive power. Given that the QSAR models we used to compute the QAFFP were selected on the basis of data availability alone, we anticipate better modeling results for QAFFP generated using more diverse and biologically meaningful targets. Data sets and Python code are publicly available at.
Název v anglickém jazyce
QSAR-derived affinity fingerprints (part 2): modeling performance for potency prediction
Popis výsledku anglicky
Affinity fingerprints report the activity of small molecules across a set of assays, and thus permit to gather information about the bioactivities of structurally dissimilar compounds, where models based on chemical structure alone are often limited, and model complex biological endpoints, such as human toxicity and in vitro cancer cell line sensitivity. Here, we propose to model in vitro compound activity using computationally predicted bioactivity profiles as compound descriptors. To this aim, we apply and validate a framework for the calculation of QSAR-derived affinity fingerprints (QAFFP) using a set of 1360 QSAR models generated using K-i, K-d, IC50 and EC50 data from ChEMBL database. QAFFP thus represent a method to encode and relate compounds on the basis of their similarity in bioactivity space. To benchmark the predictive power of QAFFP we assembled IC50 data from ChEMBL database for 18 diverse cancer cell lines widely used in preclinical drug discovery, and 25 diverse protein target data sets. This study complements part 1 where the performance of QAFFP in similarity searching, scaffold hopping, and bioactivity classification is evaluated. Despite being inherently noisy, we show that using QAFFP as descriptors leads to errors in prediction on the test set in the similar to 0.65-0.95 pIC(50) units range, which are comparable to the estimated uncertainty of bioactivity data in ChEMBL (0.76-1.00 pIC(50) units). We find that the predictive power of QAFFP is slightly worse than that of Morgan2 fingerprints and 1D and 2D physicochemical descriptors, with an effect size in the 0.02-0.08 pIC(50) units range. Including QSAR models with low predictive power in the generation of QAFFP does not lead to improved predictive power. Given that the QSAR models we used to compute the QAFFP were selected on the basis of data availability alone, we anticipate better modeling results for QAFFP generated using more diverse and biologically meaningful targets. Data sets and Python code are publicly available at.

Klasifikace

Druh
J<sub>imp</sub> - Článek v periodiku v databázi Web of Science
CEP obor
—
OECD FORD obor
10201 - Computer sciences, information science, bioinformathics (hardware development to be 2.2, social aspect to be 5.8)

Návaznosti výsledku

Projekt
<a href="/cs/project/LM2015063" target="_blank" >LM2015063: Národní infrastruktura chemické biologie</a><br>
Návaznosti
I - Institucionalni podpora na dlouhodoby koncepcni rozvoj vyzkumne organizace

Ostatní

Rok uplatnění
2020
Kód důvěrnosti údajů
S - Úplné a pravdivé údaje o projektu nepodléhají ochraně podle zvláštních právních předpisů

Údaje specifické pro druh výsledku

Název periodika
Journal of Cheminformatics
ISSN
1758-2946
e-ISSN
—
Svazek periodika
12
Číslo periodika v rámci svazku
1
Stát vydavatele periodika
GB - Spojené království Velké Británie a Severního Irska
Počet stran výsledku
17
Strana od-do
41
Kód UT WoS článku
000549151400001
EID výsledku v databázi Scopus
—

Podobné výsledky(10)

QSAR-derived affinity fingerprints (part 1): fingerprint construction and modeling performance for similarity searching, bioactivity classification and scaffold hopping QSAR Modeling Based on Conformation Ensembles Using a Multi-Instance Learning Approach QSAR – MODELOVÁNÍ KVANTITATIVNÍCH VZTAHŮ MEZI STRUKTUROU A AKTIVITOU CHEMICKÝCH LÁTEK

Co hledáte?

Rychlé hledání

Chytré vyhledávání

QSAR-derived affinity fingerprints (part 2): modeling performance for potency prediction

Identifikátory výsledku

Alternativní jazyky

Klasifikace

Návaznosti výsledku

Ostatní

Údaje specifické pro druh výsledku

Podobné výsledky(10)

Co hledáte?

Rychlé hledání

Chytré vyhledávání

Popis výsledku

Identifikátory výsledku

Identifikátory výsledku

Alternativní jazyky

Alternativní jazyky

Klasifikace

Klasifikace

Návaznosti výsledku

Návaznosti výsledku

Ostatní

Ostatní

Údaje specifické pro druh výsledku

Údaje specifické pro druh výsledku

Podobné výsledky(10)