Polymer informatics at-scale with multitask graph neural networks
arXiv:2209.13557 · doi:10.1021/acs.chemmater.2c02991
Abstract
Artificial intelligence-based methods are becoming increasingly effective at screening libraries of polymers down to a selection that is manageable for experimental inquiry. The vast majority of presently adopted approaches for polymer screening rely on handcrafted chemostructural features extracted from polymer repeat units -- a burdensome task as polymer libraries, which approximate the polymer chemical search space, progressively grow over time. Here, we demonstrate that directly "machine-learning" important features from a polymer repeat unit is a cheap and viable alternative to extracting expensive features by hand. Our approach -- based on graph neural networks, multitask learning, and other advanced deep learning techniques -- speeds up feature extraction by one to two orders of magnitude relative to presently adopted handcrafted methods without compromising model accuracy for a variety of polymer property prediction tasks. We anticipate that our approach, which unlocks the screening of truly massive polymer libraries at scale, will enable more sophisticated and large scale screening technologies in the field of polymer informatics.
References in corpus (3)
Cited by in corpus (8)
- polyBERT: A chemical language model to enable fully machine-driven ultrafast polymer informatics
- A Physics-Enforced Neural Network to Predict Polymer Melt Viscosity
- Polymer Composites Informatics for Flammability, Thermal, Mechanical and Electrical Property Predictions
- Multimodal machine learning with large language embedding model for polymer property prediction
- Benchmarking Large Language Models for Polymer Property Predictions
- Accelerating materials discovery for polymer solar cells: Data-driven insights enabled by natural language processing
- Superconductor discovery in the emerging paradigm of Materials Informatics
- AI-Driven Design of poly(ethylene terephthalate)-replacement copolymers