The generated database GDB17 enumerates 166.4 billion possible molecules up to 17 atoms of C, N, O, S and halogens following simple chemical stability and synthetic feasibility rules, however medicinal chemistry criteria are not taken into account. Here we applied rules inspired by medicinal chemistry to exclude problematic functional groups and complex molecules from GDB17, and sampled the resulting subset uniformly across molecular size, stereochemistry and polarity to form GDBMedChem as a compact collection of 10 million small molecules. This collection has reduced complexity and better synthetic accessibility than the entire GDB17 but retains higher sp 3 -carbon fraction and natural product likeness scores compared to known drugs. GDBMedChem molecules are more diverse and very different from known molecules in terms of substructures and represent an unprecedented source of diversity for drug design. GDBMedChem is available for 3D-visualization, similarity searching and for download at http://gdb.unibe.ch.
Keywords: chemical space; drug design; medicinal chemistry; small molecules; virtual screening.
© 2019 Wiley-VCH Verlag GmbH & Co. KGaA, Weinheim.
Bühlmann S, Reymond JL. Bühlmann S, et al. Front Chem. 2020 Feb 4;8:46. doi: 10.3389/fchem.2020.00046. eCollection 2020. Front Chem. 2020. PMID: 32117874 Free PMC article.
Ruddigkeit L, Blum LC, Reymond JL. Ruddigkeit L, et al. J Chem Inf Model. 2013 Jan 28;53(1):56-65. doi: 10.1021/ci300535x. Epub 2013 Jan 9. J Chem Inf Model. 2013. PMID: 23259841
Blum LC, Reymond JL. Blum LC, et al. J Am Chem Soc. 2009 Jul 1;131(25):8732-3. doi: 10.1021/ja902302h. J Am Chem Soc. 2009. PMID: 19505099
Zuegg J, Cooper MA. Zuegg J, et al. Curr Top Med Chem. 2012;12(14):1500-13. doi: 10.2174/156802612802652466. Curr Top Med Chem. 2012. PMID: 22827520 Review.
López-Vallejo F, Giulianotti MA, Houghten RA, Medina-Franco JL. López-Vallejo F, et al. Drug Discov Today. 2012 Jul;17(13-14):718-26. doi: 10.1016/j.drudis.2012.04.001. Epub 2012 Apr 10. Drug Discov Today. 2012. PMID: 22515962 Review.
Chen S, Jung Y. Chen S, et al. J Cheminform. 2024 Jul 23;16(1):83. doi: 10.1186/s13321-024-00879-0. J Cheminform. 2024. PMID: 39044299 Free PMC article.
Saigiridharan L, Hassen AK, Lai H, Torren-Peraire P, Engkvist O, Genheden S. Saigiridharan L, et al. J Cheminform. 2024 May 23;16(1):57. doi: 10.1186/s13321-024-00860-x. J Cheminform. 2024. PMID: 38778382 Free PMC article.
Pujol-Giménez J, Poirier M, Bühlmann S, Schuppisser C, Bhardwaj R, Awale M, Visini R, Javor S, Hediger MA, Reymond JL. Pujol-Giménez J, et al. ChemMedChem. 2021 Nov 5;16(21):3306-3314. doi: 10.1002/cmdc.202100467. Epub 2021 Aug 31. ChemMedChem. 2021. PMID: 34309203 Free PMC article.
Thakkar A, Chadimová V, Bjerrum EJ, Engkvist O, Reymond JL. Thakkar A, et al. Chem Sci. 2021 Jan 22;12(9):3339-3349. doi: 10.1039/d0sc05401a. Chem Sci. 2021. PMID: 34164104 Free PMC article.
Poirier M, Pujol-Giménez J, Manatschal C, Bühlmann S, Embaby A, Javor S, Hediger MA, Reymond JL. Poirier M, et al. RSC Med Chem. 2020 Jun 2;11(9):1023-1031. doi: 10.1039/d0md00085j. eCollection 2020 Sep 1. RSC Med Chem. 2020. PMID: 33479694 Free PMC article.