Back to Search
Start Over
QSAR-derived affinity fingerprints (part 1): fingerprint construction and modeling performance for similarity searching, bioactivity classification and scaffold hopping
- Source :
- Journal of Cheminformatics, 12, 39, Journal of Cheminformatics, Journal of Cheminformatics, Vol 12, Iss 1, Pp 1-16 (2020)
- Publication Year :
- 2020
-
Abstract
- Funder: FP7 People: Marie-Curie Actions; doi: http://dx.doi.org/10.13039/100011264; Grant(s): 238701, 238701<br />An affinity fingerprint is the vector consisting of compound’s affinity or potency against the reference panel of protein targets. Here, we present the QAFFP fingerprint, 440 elements long in silico QSAR-based affinity fingerprint, components of which are predicted by Random Forest regression models trained on bioactivity data from the ChEMBL database. Both real-valued (rv-QAFFP) and binary (b-QAFFP) versions of the QAFFP fingerprint were implemented and their performance in similarity searching, biological activity classification and scaffold hopping was assessed and compared to that of the 1024 bits long Morgan2 fingerprint (the RDKit implementation of the ECFP4 fingerprint). In both similarity searching and biological activity classification, the QAFFP fingerprint yields retrieval rates, measured by AUC (~ 0.65 and ~ 0.70 for similarity searching depending on data sets, and ~ 0.85 for classification) and EF5 (~ 4.67 and ~ 5.82 for similarity searching depending on data sets, and ~ 2.10 for classification), comparable to that of the Morgan2 fingerprint (similarity searching AUC of ~ 0.57 and ~ 0.66, and EF5 of ~ 4.09 and ~ 6.41, depending on data sets, classification AUC of ~ 0.87, and EF5 of ~ 2.16). However, the QAFFP fingerprint outperforms the Morgan2 fingerprint in scaffold hopping as it is able to retrieve 1146 out of existing 1749 scaffolds, while the Morgan2 fingerprint reveals only 864 scaffolds.
- Subjects :
- Quantitative structure–activity relationship
Computer science
In silico
Bioactivity modeling
Library and Information Sciences
Scaffold hopping
01 natural sciences
Biological fingerprint
lcsh:Chemistry
03 medical and health sciences
Similarity (network science)
Similarity searching
Research article
Physical and Theoretical Chemistry
030304 developmental biology
0303 health sciences
lcsh:T58.5-58.64
lcsh:Information technology
business.industry
QSAR
Fingerprint (computing)
Pattern recognition
chEMBL
Computer Graphics and Computer-Aided Design
0104 chemical sciences
Computer Science Applications
Random forest
010404 medicinal & biomolecular chemistry
lcsh:QD1-999
Big Data in Chemistry
Affinity fingerprint
Artificial intelligence
business
Research Article
Subjects
Details
- Language :
- English
- Database :
- OpenAIRE
- Journal :
- Journal of Cheminformatics, 12, 39, Journal of Cheminformatics, Journal of Cheminformatics, Vol 12, Iss 1, Pp 1-16 (2020)
- Accession number :
- edsair.doi.dedup.....62c3fe2e04882a74a8c2c7c7bc0110b9