Scaffolding learning: From specific to generic with large language models

7Citations
Citations of this article
36Readers
Mendeley users who have this article in their library.
Get full text

Abstract

Large language models such as ChatGPT have been shown to excel in solving complex math problems. However, they cannot solve basic arithmetic problems such as 758*639 = 484,362. This makes us ponder if LLMs have been trained to solve math and science problems in the right way. When a student learns math at school, she or he starts with arithmetic, then moves to word problems, polynomials, and calculus. Each skill she or he acquires will be used in the next stage to solve more advanced problems. In this paper we propose Scaffolding Learning for LLMs, which imitates how a student learns a subject in a step-by-step manner. For example, we first train an LLM to perform highly specific operations such as multiplication and division, and then apply such "skills"in a more generic task such as solving word problems. This is related to Curriculum Training, which trains a model on tasks following a specific order, such as training on easy tasks first and then gradually increases the difficulty. Our proposed approach goes from specific tasks to generic ones, which can be considered as a special case of Curriculum Training. Our empirical studies show that when an LLM has "mastered"a specific skill, only a small amount of training is required to teach it to apply the skill to a more generic application.

Cite

CITATION STYLE

APA

Yin, D. S., & Yin, X. (2024). Scaffolding learning: From specific to generic with large language models. PLoS ONE, 19(9 September). https://doi.org/10.1371/journal.pone.0310409

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free