OPERA models for predicting physicochemical properties and environmental fate endpoints

نویسندگان

  • Kamel Mansouri
  • Christopher M. Grulke
  • Richard S. Judson
  • Antony J. Williams
چکیده

The collection of chemical structure information and associated experimental data for quantitative structure-activity/property relationship (QSAR/QSPR) modeling is facilitated by an increasing number of public databases containing large amounts of useful data. However, the performance of QSAR models highly depends on the quality of the data and modeling methodology used. This study aims to develop robust QSAR/QSPR models for chemical properties of environmental interest that can be used for regulatory purposes. This study primarily uses data from the publicly available PHYSPROP database consisting of a set of 13 common physicochemical and environmental fate properties. These datasets have undergone extensive curation using an automated workflow to select only high-quality data, and the chemical structures were standardized prior to calculation of the molecular descriptors. The modeling procedure was developed based on the five Organization for Economic Cooperation and Development (OECD) principles for QSAR models. A weighted k-nearest neighbor approach was adopted using a minimum number of required descriptors calculated using PaDEL, an open-source software. The genetic algorithms selected only the most pertinent and mechanistically interpretable descriptors (2-15, with an average of 11 descriptors). The sizes of the modeled datasets varied from 150 chemicals for biodegradability half-life to 14,050 chemicals for logP, with an average of 3222 chemicals across all endpoints. The optimal models were built on randomly selected training sets (75%) and validated using fivefold cross-validation (CV) and test sets (25%). The CV Q2 of the models varied from 0.72 to 0.95, with an average of 0.86 and an R2 test value from 0.71 to 0.96, with an average of 0.82. Modeling and performance details are described in QSAR model reporting format and were validated by the European Commission's Joint Research Center to be OECD compliant. All models are freely available as an open-source, command-line application called OPEn structure-activity/property Relationship App (OPERA). OPERA models were applied to more than 750,000 chemicals to produce freely available predicted data on the U.S. Environmental Protection Agency's CompTox Chemistry Dashboard.

برای دانلود رایگان متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

BiodegHL: Biodegradation half-life prediction from OPERA (OPEn saR App) models

BiodegHL: Biodegradation half-life prediction from OPERA (OPEn saR App) models. 1.2.Other related models: No related models 1.3.Software coding the model: OPERA V1.02 OPERA (OPEn (quantitative) structure-activity Relationship Application) is a standalone free and open source command line application. It provides a suite of QSAR models to predict physicochemical properties and environmental fate...

متن کامل

MP: Melting point prediction from OPERA (OPEn saR App) models

MP: Melting point prediction from OPERA (OPEn saR App) models. 1.2.Other related models: No related models 1.3.Software coding the model: OPERA V1.02 OPERA (OPEn (quantitative) structure-activity Relationship Application) is a standalone free and open source command line application. It provides a suite of QSAR models to predict physicochemical properties and environmental fate of organic chemi...

متن کامل

VP: Vapor pressure prediction from OPERA (OPEn saR App) models

VP: Vapor pressure prediction from OPERA (OPEn saR App) models. 1.2.Other related models: No related models 1.3.Software coding the model: OPERA V1.02 OPERA (OPEn (quantitative) structure-activity Relationship Application) is a standalone free and open source command line application. It provides a suite of QSAR models to predict physicochemical properties and environmental fate of organic chem...

متن کامل

BP: Boiling point prediction from OPERA (OPEn saR App) models

BP: Boiling point prediction from OPERA (OPEn saR App) models. 1.2.Other related models: No related models 1.3.Software coding the model: OPERA V1.02 OPERA (OPEn (quantitative) structure-activity Relationship Application) is a standalone free and open source command line application. It provides a suite of QSAR models to predict physicochemical properties and environmental fate of organic chemi...

متن کامل

LogP: Octanol-water partition coefficient prediction from the OPERA (OPEn saR App) models

LogP: Octanol-water partition coefficient prediction from the OPERA (OPEn saR App) models. 1.2.Other related models: No related models 1.3.Software coding the model: OPERA V1.02 OPERA (OPEn (quantitative) structure-activity Relationship Application) is a standalone free and open source command line application. It provides a suite of QSAR models to predict physicochemical properties and environ...

متن کامل

ذخیره در منابع من


  با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

عنوان ژورنال:

دوره 10  شماره 

صفحات  -

تاریخ انتشار 2018