Hugging Face Trending Papers

PolymerGPT: Multi-property Optimization with a Decoder-Based GPT Model for Generative Polymer Design

Polymer property prediction and inverse generative design targeting desired properties are two crucial tasks in machine learning-assisted polymer design. While the former has received considerable attention, there have been limited methods developed for the latter.

arXiv AI
Sep 3

HiPoly: a hierarchical polymer-native AI framework for property prediction and generative design

HiPoly is a polymer-native AI framework that uses a three-level hierarchical graph architecture built on the G2RINS representation to process complete polymer descriptions. It encodes stochastic inter-monomer connectivity, composition, and molecular weight directly within its architecture, enabling end-to-end workflows from experimental data to property prediction, generative design, and physics-based validation. The framework achieves state-of-the-art accuracy for thermophysical properties of multi-component polymer systems and demonstrates generative design by discovering sustainable, PFAS-free alternatives with target surface-energy properties.

By Ge Sun, Gervasio Zaldivar, Yuan Tian, Gustavo Perez Lemus, Juhae Park, Dasha Safarian, Ming Han, Juan J. de Pablo
arXiv AI
Sep 24

An open benchmark for machine learning-based polymer property prediction

The article introduces PolyBench26, an open benchmark dataset for polymer property prediction that contains nearly 250,000 datapoints covering eight physical properties from experimental, DFT, and MD sources. It supports four evaluation tasks—property prediction, dataset-size scaling, repeat‑unit complexity, and transfer learning—across homopolymers and various copolymer architectures. The study compares language, graph, and descriptor models, finding graph-based approaches achieve the lowest errors and maintain robustness across training sizes and repeat‑unit complexity.

By Robert W. Learsch, Nicholas Liesen, Daniel S. Levine, Anna M. Hiszpanski, Evan R. Antoniuk
arXiv Machine Learning
Sep 16

Molecular representation shapes the balance between target fidelity and exploration in flow based polymer generation

The paper introduces PolyLatentFlow, a continuous‑time flow‑matching framework for polymer generation, and LlamaUni, a multimodal representation that fuses polymer sequences with 3D structural data. In unconditional generation, the combination yields the highest number of valid, novel candidates while preserving diversity, and in conditional settings it systematically shifts property distributions across a 200 °C target range. Across multi‑property tasks, the representation choice affects validity, training‑set replay, and structural proximity, with PolyLatentFlow + LlamaUni achieving the best balance of high validity, low replay, and high target hit yield.

By Tianren Zhang
arXiv AI
Aug 18

Empowering Polymeric Materials Discovery by Artificial Intelligence

arXiv:2606. 20753v2 Announce Type: replace-cross Abstract: Polymeric materials underpin modern technologies spanning energy storage, microelectronics, healthcare and sustainable manufacturing.

By Chenyao Ma, Linda Zhang, Yuheng Chen, Wei Du, Shangwen Fang, Zihao Jiang, Chuanyu Liu, Xinyu Ma, Rui Su, Gang Wang, Muyao Yu, Dong Zhong, Jie Zhu, Weibo Gong, Huan Gu, Limin Li, Chen Shen, Rui Wu, Zhenghao Wu, Kan Xu, Min Zhou, Donglin He, Xiayun Huang, Shan Jiang, Pengfei Ou, Jiayu Peng, Yuwei Zhang, Jie Zhao, Di Zhang, Piao Ma, Zhenghao Li, Hao Li
arXiv Machine Learning
Jun 2

Towards Automated Discovery: A Review of Generative Models, Multimodal Learning and Closed-Loop Workflows in Inverse Materials Design

arXiv:2606. 02507v1 Announce Type: cross Abstract: Inverse materials design is shifting materials discovery from forward prediction to targeted proposal of candidates that satisfy objectives under physical constraints.

By Anand Babu, Rog\'erio Almeida Gouv\^ea, Gian-Marco Rignanese
arXiv Machine Learning
Sep 17

Active Learning Enables Generation of Molecules that Advance the Known Pareto Front

The paper presents a closed‑loop molecule generation pipeline that iteratively retrains on new quantum‑chemical simulation data, overcoming limitations of static generative models. This approach produces molecules whose properties extend up to 0.44 standard deviations beyond the training set and improves out‑of‑distribution classification accuracy by 79%. By conditioning on thermodynamic stability during the loop, the method yields a 3.5‑fold increase in the proportion of stable, potentially synthesizable molecules.

By Evan R. Antoniuk, Peggy Li, Nathan Keilbart, Stephen Weitzner, Bhavya Kailkhura, Anna M. Hiszpanski
arXiv Machine Learning
Jun 17

Toward Controllable Catalyst Inverse Design via Large-Scale Autoregressive Pretraining

arXiv:2606. 17445v1 Announce Type: new Abstract: Inverse design of heterogeneous catalysts remains challenging because catalyst surfaces exhibit substantial structural complexity with coupled surface-adsorbate interactions across a vast chemical space that is difficult to explore efficiently through conventional screening alone.

By Dong Hyeon Mok, Jonggeol Na, Seoin Back