首页    期刊浏览 2024年09月21日 星期六
登录注册

文章基本信息

  • 标题:SYBA: Bayesian estimation of synthetic accessibility of organic compounds
  • 本地全文:下载
  • 作者:Milan Voršilák ; Michal Kolář ; Ivan Čmelo
  • 期刊名称:Journal of Cheminformatics
  • 印刷版ISSN:1758-2946
  • 电子版ISSN:1758-2946
  • 出版年度:2020
  • 卷号:12
  • 期号:1
  • 页码:1-13
  • DOI:10.1186/s13321-020-00439-2
  • 出版社:BioMed Central
  • 摘要:SYBA (SYnthetic Bayesian Accessibility) is a fragment-based method for the rapid classification of organic compounds as easy- (ES) or hard-to-synthesize (HS). It is based on a Bernoulli naïve Bayes classifier that is used to assign SYBA score contributions to individual fragments based on their frequencies in the database of ES and HS molecules. SYBA was trained on ES molecules available in the ZINC15 database and on HS molecules generated by the Nonpher methodology. SYBA was compared with a random forest, that was utilized as a baseline method, as well as with other two methods for synthetic accessibility assessment: SAScore and SCScore. When used with their suggested thresholds, SYBA improves over random forest classification, albeit marginally, and outperforms SAScore and SCScore. However, upon the optimization of SAScore threshold (that changes from 6.0 to – 4.5), SAScore yields similar results as SYBA. Because SYBA is based merely on fragment contributions, it can be used for the analysis of the contribution of individual molecular parts to compound synthetic accessibility. SYBA is publicly available at https://github.com/lich-uct/syba under the GNU General Public License.
  • 关键词:Synthetic accessibility;Bayesian analysis;Bernoulli naïve Bayes
国家哲学社会科学文献中心版权所有