Visualization of the Chemical Space in Drug Discovery
José L. Medina‐Franco, Karina Martínez‐Mayorga, Marc A. Giulianotti, Richard A. Houghten, Clemencia Pinilla
Abstract
José L. Medina‐Franco, Karina Martínez‐Mayorga, Marc A. Giulianotti, Richard A. Houghten, Clemencia Pinilla
Abstract
Chemical space has become a key concept in drug discovery. The continued growth in the number of molecules available raises the question regarding how many compounds may exist and which ones have the potential to become drugs. Analysis and visualization of the chemical space covered by public, commercial, in-house and virtual compound collections have found multiple applications in diversity analysis, in silico property profiling, data mining, virtual screening, library design, prioritization in screening campaigns, and acquisition of compound collections, among others. This review covers several techniques, computational programs and approaches that have been developed to visualize, navigate and study the chemical space of molecular databases. Techniques developed in our group are presented including a quantitative assessment of the multi-fusion similarity maps. Additionally an application of 3D-similarity, based on the overlay of chemical structures, to represent the chemical space is introduced. Several comparisons of the chemical space covered by compound collections from different sources such as combinatorial libraries, drugs and natural products, or directed to specific therapeutic areas are also discussed. Keywords: Chemoinformatics, combinatorial libraries, data-driven analysis, data mining, molecular diversity, multi-fusion similarity maps, structure-activity relationships, virtual screening
OpenAlex reports 169 citations for this work. Citation counts describe recorded attention and do not establish research quality.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
Chemical space has become a key concept in drug discovery. The continued growth in the number of molecules available raises the question regarding how many compounds may exist and which ones have the potential to become drugs. Analysis and visualization of the chemical space covered by public, commercial, in-house and virtual compound collections have found multiple applications in diversity analysis, in silico property profiling, data mining, virtual screening, library design, prioritization in screening campaigns, and acquisition of compound collections, among others. This review covers several techniques, computational programs and approaches that have been developed to visualize, navigate and study the chemical space of molecular databases. Techniques developed in our group are presented including a quantitative assessment of the multi-fusion similarity maps. Additionally an application of 3D-similarity, based on the overlay of chemical structures, to represent the chemical space is introduced. Several comparisons of the chemical space covered by compound collections from different sources such as combinatorial libraries, drugs and natural products, or directed to specific therapeutic areas are also discussed. Keywords: Chemoinformatics, combinatorial libraries, data-driven analysis, data mining, molecular diversity, multi-fusion similarity maps, structure-activity relationships, virtual screening
Key concepts: Chemical space, Cheminformatics, Virtual screening, Computer science, Visualization, Chemical database, Chemical similarity, Drug discovery