Evaluating In-Context Learning in Large Language Models for Molecular Property Regression

0Citations
Citations of this article
5Readers
Mendeley users who have this article in their library.
Get full text

Abstract

Large language models (LLMs) demonstrate strong performance in natural language tasks, but their capacity for genuine in-context learning (ICL) in scientific regression remains unclear. We systematically assessed seven LLMs on molecular property prediction using a controlled framework of 56 transformed tasks that isolate shortcut learning and are designed to induce functional out-of-distribution (OOD) behavior. LLMs performed nearly perfectly on raw molecular weight prediction via shortcut cues but deteriorated under nonlinear transformations, whereas machine learning (ML) baselines showed greater robustness, yielding a performance crossover. Meta-analysis revealed that distributional descriptors and structure–activity landscape indices (SALI) predict task favorability, providing a framework for selecting between LLM- and ML-based approaches in chemistry.

Cite

CITATION STYLE

APA

Joe, C. Y., Song, K., & Chang, R. (2026). Evaluating In-Context Learning in Large Language Models for Molecular Property Regression. Journal of Computational Chemistry, 47(2). https://doi.org/10.1002/jcc.70308

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free