Structural Diversity of Organic Chemistry. A Scaffold Analysis of the CAS Registry
Author(s) -
Alan H. Lipkus,
Qiong Yuan,
Karen A. Lucas,
Susan A. Funk,
William F. Bartelt,
Roger J. Schenck,
Anthony J. Trippe
Publication year - 2008
Publication title -
the journal of organic chemistry
Language(s) - English
Resource type - Journals
SCImago Journal Rank - 1.2
H-Index - 228
eISSN - 1520-6904
pISSN - 0022-3263
DOI - 10.1021/jo8001276
Subject(s) - scaffold , minification , computer science , ring (chemistry) , key (lock) , organic molecules , chemistry , combinatorial chemistry , distribution (mathematics) , molecule , biochemical engineering , theoretical computer science , organic chemistry , mathematics , engineering , computer security , database , programming language , mathematical analysis
By analyzing the scaffold content of the CAS Registry, we attempt to characterize in a comprehensive way the structural diversity of organic chemistry. The scaffold of a molecule is taken to be its framework, defined as all its ring systems and all the linkers that connect them. Framework data from more than 24 million organic compounds is analyzed. The distribution of frameworks among compounds is found to be top-heavy, i.e., a small percentage of frameworks occur in a large percentage of compounds. When frameworks are analyzed at the graph level, an even more top-heavy distribution is found: half of the compounds can be described by only 143 framework shapes. The most significant finding is that the framework distribution conforms almost exactly to a power law. This suggests that the more often a framework has been used as the basis for a compound, the more likely it is to be used in another compound. This may be explained by the cost of synthesis: making a new derivative of a framework is probably less costly if many other derivatives are known. We believe this power law is evidence that the minimization of synthetic cost has been a key factor in shaping the known universe of organic chemistry.
Accelerating Research
Robert Robinson Avenue,
Oxford Science Park, Oxford
OX4 4GP, United Kingdom
Address
John Eccles HouseRobert Robinson Avenue,
Oxford Science Park, Oxford
OX4 4GP, United Kingdom