February 6, 2024

GPT-3 transforms chemical research

by Ecole Polytechnique Federale de Lausanne

Artificial intelligence is growing into a pivotal tool in chemical research, offering novel methods to tackle complex challenges that traditional approaches struggle with. One subtype of artificial intelligence that has seen increasing use in chemistry is machine learning, which uses algorithms and statistical models to make decisions based on data and perform tasks that it has not been explicitly programmed for.

However, to make reliable predictions, machine learning also demands large amounts of data, which isn't always available in chemical research. Small chemical datasets simply do not provide enough information for these algorithms to train on, which limits their effectiveness.

Scientists, in the team of Berend Smit at EPFL, have found a solution in large language models such as GPT-3. Those models are pre-trained on massive amounts of texts, and are known for their broad capabilities in understanding and generating human-like text. GPT-3 forms the basis of the more popular artificial intelligence ChatGPT.

The study, published in Nature Machine Intelligence, unveils a novel approach that significantly simplifies chemical analysis using artificial intelligence. Contrary to initial skepticism, the method doesn't directly ask GPT-3 chemical questions.

"GPT-3 has not seen most of the chemical literature, so if we ask ChatGPT a chemical question, the answers are typically limited to what one can find on Wikipedia," says Kevin Jablonka, the study's lead researcher.

"Instead, we fine-tune GPT-3 with a small data set converted into questions and answers, creating a new model capable of providing accurate chemical insights."

This process involves feeding GPT-3 a curated list of Q&As. "For example, for high-entropy alloys, it is important to know whether an alloy occurs in a single phase or has multiple phases," says Smit. "The curated list of Q&As are of the type: Q= 'Is the (name of the high entropy alloy) single phase?' A= 'Yes/No.'"

He continues, "In the literature, we have found many alloys of which the answer is known, and we used this data to fine-tune GPT-3. What we get back is a refined AI model that is trained to only answer this question with a yes or no."

In tests, the model, trained with relatively few Q&As, correctly answered over 95% of very diverse chemical problems, often surpassing the accuracy of state-of-the-art machine-learning models. "The point is that this is as easy as doing a literature search, which works for many chemical problems," says Smit.

One of the most striking aspects of this study is its simplicity and speed. Traditional machine learning models require months to develop and demand extensive knowledge. In contrast, the approach developed by Jablonka takes five minutes and requires zero knowledge.

The implications of the study are profound. It introduces a method as easy as conducting a literature search, applicable to various chemical problems. The ability to formulate questions like "Is the yield of a [chemical] made with this (recipe) high?" and receive accurate answers can revolutionize how chemical research is planned and carried out.

In the paper, the authors say, "Next to a literature search, querying a foundational model (e.g., GPT-3,4) might become a routine way to bootstrap a project by leveraging the collective knowledge encoded in these foundational models." Or, as Smit succinctly puts it, "This is going to change the way we do chemistry."

More information: Kevin Maik Jablonka, Is GPT all you need for low-data discovery in chemistry?, Nature Machine Intelligence (2024). DOI: 10.1038/s42256-023-00788-1

Journal information: Nature Machine Intelligence

Provided by Ecole Polytechnique Federale de Lausanne

Citation: GPT-3 transforms chemical research (2024, February 6) retrieved 7 March 2024 from https://phys.org/news/2024-02-gpt-chemical.html

This document is subject to copyright. Apart from any fair dealing for the purpose of private study or research, no part may be reproduced without the written permission. The content is provided for information purposes only.

Explore further

Study explores the scaling of deep learning models for chemistry research

69 shares

Feedback to editors

GPT-3 transforms chemical research

AI makes a rendezvous in space

Vitamin A may play a central role in stem cell biology and wound repair

Researchers test curcumin nanoemulsion for treatment of intestinal inflammation

New technique may help scientists stave off coral reef collapse

New study reveals which animals are most vulnerable to extinction due to climate change

Researchers reveal how a virus hijacks insect sperm: May help control disease vectors and pests

Geologists find that low-relief mountain ranges are the largest carbon sinks

Often seen, never studied: First characterization of a key postsynaptic protein

New study finds the malaria parasite generates genetic diversity using an evolutionary 'copy-paste' tactic

Earth's earliest forest revealed in Somerset fossils

Relevant PhysicsForums posts

New Insight into the Chemistry of Solvents

Quantum hybridized orbitals

BET SA and Microporous Materials - equipment/software for MOFs

Help please with SI-ATRP (polymerization) of polyacrylamide

Debunking the Hypervalence Theory: The Truth About Bonds in HClO4

Central atoms in Lewis structures: basic question

Study explores the scaling of deep learning models for chemistry research

AI can help forecast air quality, but freak events like 2023's summer of wildfire smoke require traditional methods too

New AI model transforms understanding of metal-organic frameworks

Machine learning cracks the oxidation states of crystal structures

Machine learning predicts heat capacities of metal-organic frameworks

Using machine learning to forecast amine emissions

Scientists develop new machine learning method for modeling chemical reactions

Deciphering catalysts: Unveiling structure-activity correlations

Scientists reveal molecular mysteries to control silica scaling in water treatment

Chemists break barriers and open up super-resolution molecule mass analysis

Metal-organic framework research makes key advance toward removing pesticide from groundwater

Earth-abundant iron catalysis enables access to valuable dialkylated compounds

Medical Xpress

Tech Xplore

Science X

GPT-3 transforms chemical research

AI makes a rendezvous in space

Vitamin A may play a central role in stem cell biology and wound repair

Researchers test curcumin nanoemulsion for treatment of intestinal inflammation

New technique may help scientists stave off coral reef collapse

New study reveals which animals are most vulnerable to extinction due to climate change

Researchers reveal how a virus hijacks insect sperm: May help control disease vectors and pests

Geologists find that low-relief mountain ranges are the largest carbon sinks

Often seen, never studied: First characterization of a key postsynaptic protein

New study finds the malaria parasite generates genetic diversity using an evolutionary 'copy-paste' tactic

Earth's earliest forest revealed in Somerset fossils

Relevant PhysicsForums posts

Related Stories

Study explores the scaling of deep learning models for chemistry research

AI can help forecast air quality, but freak events like 2023's summer of wildfire smoke require traditional methods too

New AI model transforms understanding of metal-organic frameworks

Machine learning cracks the oxidation states of crystal structures

Machine learning predicts heat capacities of metal-organic frameworks

Using machine learning to forecast amine emissions

Recommended for you

Scientists develop new machine learning method for modeling chemical reactions

Deciphering catalysts: Unveiling structure-activity correlations

Scientists reveal molecular mysteries to control silica scaling in water treatment

Chemists break barriers and open up super-resolution molecule mass analysis

Metal-organic framework research makes key advance toward removing pesticide from groundwater

Earth-abundant iron catalysis enables access to valuable dialkylated compounds

Newsletter sign up

Donate and enjoy an ad-free experience