How Well Can Frontier Large Language Models Generate Structures? High Quality Prediction of Molecular Geometries with Help from Fine-Tuning
Joseph M. Cavanagh, Jonathan B. Arnold, Giovanni Battista Alteri, Andrew Gritsevskiy, Teresa Head-Gordon
Abstract
The power of Large Language Models (LLMs) has led us to investigate how they might be fine-tuned for learning the "language of molecular geometry". The fine-tuning of the LLMs using Cartesian coordinates or Z-matrices provides an extremely simple method for accurately predicting equilibrium structures and diverse sets of conformers of small organic and drug-like molecules with excellent accuracy and outperforming specialized deep learning models. While the most common representation of molecular geometries are Cartesian coordinates performs adequately, we find that the inherent invariances and relational nature of geometries represented as Z-matrices provides a better grammar for LLM adaptation. Finally, we show that enhancing an LLMs capabilities for robust prediction of small molecule geometries still retains nearly all of its pre-trained language abilities by randomly mixing in small quantities of natural language prompt-response pairs into the fine-tuning.
Create a lesson
Related papers
Real-Time Emergence of Charge-Transfer-to-Solvent States from Core Excitation
Jiří Suchan, B. Scott Fales, Benjamin G. Levine et al.
ElemCo.jl: A Julia package for electron-correlation methods
Daniel Kats, Charlotte Rickert, Thomas Schraivogel et al.
Franson-Interferometric Bounds on Entangled Two-Photon Absorption
Albin Hedse, Sankaran Ramesh, Luis Matheis et al.
The off-diagonal low rank property: new opportunities for low-scaling computational chemistry methods
Zikuan Wang
Core-valence double ionization of SF6 involving S2p, F1s and S1s inner shells
Veronica Daver Ideböhn, Daniel M. Pereira, Lucas M. Cornetta et al.
Benchmark of Multi-Channel Dyson Equation and Algebraic Diagrammatic Construction Methods for molecules
Mike Keizer, Stefano Paggi, J. Arjan Berger et al.