首页    期刊浏览 2024年12月14日 星期六
登录注册

文章基本信息

  • 标题:Extended similarity indices: the benefits of comparing more than two objects simultaneously. Part 1: Theory and characteristics †
  • 本地全文:下载
  • 作者:Ramón Alain Miranda-Quintana ; Dávid Bajusz ; Anita Rácz
  • 期刊名称:Journal of Cheminformatics
  • 印刷版ISSN:1758-2946
  • 电子版ISSN:1758-2946
  • 出版年度:2021
  • 卷号:13
  • 期号:1
  • 页码:1-18
  • DOI:10.1186/s13321-021-00505-3
  • 出版社:BioMed Central
  • 摘要:Quantification of the similarity of objects is a key concept in many areas of computational science. This includes cheminformatics, where molecular similarity is usually quantified based on binary fingerprints. While there is a wide selection of available molecular representations and similarity metrics, there were no previous efforts to extend the computational framework of similarity calculations to the simultaneous comparison of more than two objects (molecules) at the same time. The present study bridges this gap, by introducing a straightforward computational framework for comparing multiple objects at the same time and providing extended formulas for as many similarity metrics as possible. In the binary case (i.e. when comparing two molecules pairwise) these are naturally reduced to their well-known formulas. We provide a detailed analysis on the effects of various parameters on the similarity values calculated by the extended formulas. The extended similarity indices are entirely general and do not depend on the fingerprints used. Two types of variance analysis (ANOVA) help to understand the main features of the indices: (i) ANOVA of mean similarity indices; (ii) ANOVA of sum of ranking differences (SRD). Practical aspects and applications of the extended similarity indices are detailed in the accompanying paper: Miranda-Quintana et al. J Cheminform. 2021. https://doi.org/10.1186/s13321-021-00504-4 . Python code for calculating the extended similarity metrics is freely available at: https://github.com/ramirandaq/MultipleComparisons .
  • 关键词:Comparisons ; Rankings ; Extended similarity indices ; Consistency ; Molecular fingerprints ; ANOVA ; Sum of ranking differences
国家哲学社会科学文献中心版权所有