Exploring activity landscapes with extended similarity : is Tanimoto enough?
© 2023 The Authors. Molecular Informatics published by Wiley-VCH GmbH..
Understanding structure-activity landscapes is essential in drug discovery. Similarly, it has been shown that the presence of activity cliffs in compound data sets can have a substantial impact not only on the design progress but also can influence the predictive ability of machine learning models. With the continued expansion of the chemical space and the currently available large and ultra-large libraries, it is imperative to implement efficient tools to analyze the activity landscape of compound data sets rapidly. The goal of this study is to show the applicability of the n-ary indices to quantify the structure-activity landscapes of large compound data sets using different types of structural representation rapidly and efficiently. We also discuss how a recently introduced medoid algorithm provides the foundation to finding optimum correlations between similarity measures and structure-activity rankings. The applicability of the n-ary indices and the medoid algorithm is shown by analyzing the activity landscape of 10 compound data sets with pharmaceutical relevance using three fingerprints of different designs, 16 extended similarity indices, and 11 coincidence thresholds.
Medienart: |
E-Artikel |
---|
Erscheinungsjahr: |
2023 |
---|---|
Erschienen: |
2023 |
Enthalten in: |
Zur Gesamtaufnahme - volume:42 |
---|---|
Enthalten in: |
Molecular informatics - 42(2023), 7 vom: 31. Juli, Seite e2300056 |
Sprache: |
Englisch |
---|
Beteiligte Personen: |
Dunn, Timothy B [VerfasserIn] |
---|
Links: |
---|
Themen: |
Chemical space |
---|
Anmerkungen: |
Date Completed 13.07.2023 Date Revised 18.07.2023 published: Print-Electronic Citation Status MEDLINE |
---|
doi: |
10.1002/minf.202300056 |
---|
funding: |
|
---|---|
Förderinstitution / Projekttitel: |
|
PPN (Katalog-ID): |
NLM357049101 |
---|
LEADER | 01000naa a22002652 4500 | ||
---|---|---|---|
001 | NLM357049101 | ||
003 | DE-627 | ||
005 | 20231226071744.0 | ||
007 | cr uuu---uuuuu | ||
008 | 231226s2023 xx |||||o 00| ||eng c | ||
024 | 7 | |a 10.1002/minf.202300056 |2 doi | |
028 | 5 | 2 | |a pubmed24n1190.xml |
035 | |a (DE-627)NLM357049101 | ||
035 | |a (NLM)37202375 | ||
040 | |a DE-627 |b ger |c DE-627 |e rakwb | ||
041 | |a eng | ||
100 | 1 | |a Dunn, Timothy B |e verfasserin |4 aut | |
245 | 1 | 0 | |a Exploring activity landscapes with extended similarity |b is Tanimoto enough? |
264 | 1 | |c 2023 | |
336 | |a Text |b txt |2 rdacontent | ||
337 | |a ƒaComputermedien |b c |2 rdamedia | ||
338 | |a ƒa Online-Ressource |b cr |2 rdacarrier | ||
500 | |a Date Completed 13.07.2023 | ||
500 | |a Date Revised 18.07.2023 | ||
500 | |a published: Print-Electronic | ||
500 | |a Citation Status MEDLINE | ||
520 | |a © 2023 The Authors. Molecular Informatics published by Wiley-VCH GmbH. | ||
520 | |a Understanding structure-activity landscapes is essential in drug discovery. Similarly, it has been shown that the presence of activity cliffs in compound data sets can have a substantial impact not only on the design progress but also can influence the predictive ability of machine learning models. With the continued expansion of the chemical space and the currently available large and ultra-large libraries, it is imperative to implement efficient tools to analyze the activity landscape of compound data sets rapidly. The goal of this study is to show the applicability of the n-ary indices to quantify the structure-activity landscapes of large compound data sets using different types of structural representation rapidly and efficiently. We also discuss how a recently introduced medoid algorithm provides the foundation to finding optimum correlations between similarity measures and structure-activity rankings. The applicability of the n-ary indices and the medoid algorithm is shown by analyzing the activity landscape of 10 compound data sets with pharmaceutical relevance using three fingerprints of different designs, 16 extended similarity indices, and 11 coincidence thresholds | ||
650 | 4 | |a Journal Article | |
650 | 4 | |a Research Support, Non-U.S. Gov't | |
650 | 4 | |a chemical space | |
650 | 4 | |a eSALI, similarity | |
650 | 4 | |a extended similarity | |
650 | 4 | |a molecular fingerprints | |
650 | 4 | |a structure-activity relationships | |
700 | 1 | |a López-López, Edgar |e verfasserin |4 aut | |
700 | 1 | |a Kim, Taewon David |e verfasserin |4 aut | |
700 | 1 | |a Medina-Franco, José L |e verfasserin |4 aut | |
700 | 1 | |a Miranda-Quintana, Ramón Alain |e verfasserin |4 aut | |
773 | 0 | 8 | |i Enthalten in |t Molecular informatics |d 2010 |g 42(2023), 7 vom: 31. Juli, Seite e2300056 |w (DE-627)NLM209791799 |x 1868-1751 |7 nnns |
773 | 1 | 8 | |g volume:42 |g year:2023 |g number:7 |g day:31 |g month:07 |g pages:e2300056 |
856 | 4 | 0 | |u http://dx.doi.org/10.1002/minf.202300056 |3 Volltext |
912 | |a GBV_USEFLAG_A | ||
912 | |a GBV_NLM | ||
951 | |a AR | ||
952 | |d 42 |j 2023 |e 7 |b 31 |c 07 |h e2300056 |