返回
. Developing Collaborative QSAR Models Without Sharing Structures
DOI:10.1021/acs.jcim.7b00315.png)
摘要
En 中文
It is widely understood that QSAR models greatly improve if more data are used. However, irrespective of model quality, once chemical structures diverge too far from the initial data set, the predictive performance of a model degrades quickly. To increase the applicability domain we need to increase the diversity of the training set. This can be achieved by combining data from diverse sources. Public data can be easily included; however, proprietary data may be more difficult to add due to intellectual property concerns. In this contribution, we will present a method for the collaborative development of linear regression models that addresses this problem. The method differs from other past approaches, because data are only shared in an aggregated form. This prohibits access to individual data points and therefore avoids the disclosure of confidential structural information. The final models are equivalent to models that were built with combined data sets.
Keyword:
ATOMIC PHYSICOCHEMICAL PARAMETERS
CALCULATING LOG P(OCT)
CHEMICAL-STRUCTURES
PREDICTION
ACCURACY
DATASET
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
5.3
论文数:
9.1K
被引数:
4.0W
机构
引用论文
Arsenic Removal from Natural Water Using Low Cost Granulated Adsorbents: A Review使用低成本颗粒状吸附剂从天然水中去除砷: 综述
Big pharma screening collections: more of the same or unique libraries? The AstraZeneca-Bayer Pharma AG case
DRUG DISCOVERY TODAY
IF7.5
Application of ALOGPS 2.1 to predict log D distribution coefficient for Pfizer proprietary compounds
Effect of the SSeCKS–TRAF6 interaction on gastrodin-mediated protection against 2,3,7,8-tetrachlorodibenzo-p-dioxin-induced astrocyte activation and neuronal death
Chemosphere
IF0

