The COMPAS Project: A Computational Database of Polycyclic Aromatic Systems.
Phase 1: cata-condensed Polybenzenoid Hydrocarbons

Alexandra Wahab; Lara Pfuderer; Eno Paenurk; Renana Gershoni-Poranne

doi:10.26434/chemrxiv-2022-2l1m9

Organic Chemistry

Search within Organic Chemistry

The COMPAS Project: A Computational Database of Polycyclic Aromatic Systems. Phase 1: cata-condensed Polybenzenoid Hydrocarbons

27 April 2022, Version 1

Working Paper

Show author details

This content is a preprint and has not undergone peer review at the time of posting.

Abstract

Chemical databases are an essential tool for data-driven investigation of structure-property relationships and design of novel functional compounds. We introduce the first phase of the COMPAS Project – a COMputational database of Polycyclic Aromatic Systems. In this phase, we have developed two datasets containing the optimized ground-state structures and a selec- tion of molecular properties of 34k and 9k cata- condensed polybenzenoid hydrocarbons (at the GFN2-xTB and B3LYP-D3BJ/def2-SVP lev- els, respectively), and have placed them in the public domain. Herein we describe the process of the dataset generation, detail the informa- tion available within the datasets, and show the fundamental features of the generated data. We analyze the correlation between the two types of computation as well as the structure- property relationships of the calculated species. The data and the insights gained from them can inform rational design of novel functional aro- matic molecules for use in, e.g., organic elec- tronics, and can provide a basis for additional data-driven machine- and deep-learning studies in chemistry.

Keywords

polycyclic aromatic hydrocarbons

database

computational chemistry

structure-property relationships

data-driven

Supplementary materials

Title

Description

Actions

Title

Supporting Information for COMPAS_Phase1

Description

General computational details, description of benchmarking procedure, histograms of data distribution, color-coded plots for all studied structural features, further analysis on D3 versus D4 corrections.

Actions

Supplementary weblinks

Title

Description

Actions

Title

Repository of COMPAS

Description

Freely accessible repository of the COMPAS database.

Actions

View

Comments

Comments are not moderated before they are posted, but they can be removed by the site moderators if they are found to be in contravention of our Commenting Policy - please read this policy before you post. Comments should be used for scholarly discussion of the content in question. You can find more information about how to use the commenting feature here .

This site is protected by reCAPTCHA and the Google Privacy Policy and Terms of Service apply.

Now Published

The COMPAS Project: A Computational Database of Polycyclic Aromatic Systems. Phase 1: cata-Condensed Polybenzenoid Hydrocarbons

Alexandra Wahab, Lara Pfuderer, Eno Paenurk, Renana Gershoni-Poranne journal article

Journal of Chemical Information and Modeling , Volume 62, Issue 16

Online publication date: Jul 26, 2022

Version History

Apr 27, 2022 Version 1

Metrics

1,207

920

Views

Downloads

License

The content is available under CC BY NC 4.0

DOI

10.26434/chemrxiv-2022-2l1m9

Funding

Branco Weiss Fellowship – Society in Science

Author’s competing interest statement

The author(s) have declared they have no conflict of interest with regard to this content

Ethics

The author(s) have declared ethics committee/IRB approval is not relevant to this content

The COMPAS Project: A Computational Database of Polycyclic Aromatic Systems. Phase 1: cata-condensed Polybenzenoid Hydrocarbons

Authors

Abstract

Keywords

Supplementary materials

Supplementary weblinks

Comments

Now Published

Version History

Metrics

License

DOI

Funding

Author’s competing interest statement

Ethics

Share