Metacognitive Capabilities of LLMs: An Exploration in Mathematical Problem Solving

Metacognitive knowledge refers to humans' intuitive knowledge of their own thinking and reasoning processes. Today's best LLMs clearly possess some reasoning processes. The paper gives evidence that they also have metacognitive knowledge, including ability to name skills and procedures to apply given a task. We explore this primarily in context of math reasoning, developing a prompt-guided interaction procedure to get a powerful LLM to assign sensible skill labels to math questions, followed by having it perform semantic clustering to obtain coarser families of skill labels. These coarse skill labels look interpretable to humans. To validate that these skill labels are meaningful and relevant to the LLM's reasoning processes we perform the following experiments. (a) We ask GPT-4 to assign skill labels to training questions in math datasets GSM8K and MATH. (b) When using an LLM to solve the test questions, we present it with the full list of skill labels and ask it to identify the skill needed. Then it is presented with randomly selected exemplar solved questions associated with that skill label. This improves accuracy on GSM8k and MATH for several strong LLMs, including code-assisted models. The methodology presented is domain-agnostic, even though this article applies it to math problems.

Solving GeneralArithmetic Word ProblemsSolving General Arithmetic Word ProblemsAre NLP Models reallyable to Solve Simple…Are NLP Models really able to Solve Simple Math Word Problems?Llama 2: Open Foundationand Fine-Tuned Chat…Llama 2: Open Foundation and Fine-Tuned Chat ModelsGemini: A Family ofHighly Capable…Gemini: A Family of Highly Capable Multimodal ModelsGPT-4 Technical ReportGPT-4 Technical ReportLLaMA: Open andEfficient Foundation…LLaMA: Open and Efficient Foundation Language ModelsPaLM 2 Technical ReportPaLM 2 Technical ReportSolving Math WordProblems by Combining…Solving Math Word Problems by Combining Language Models With Symbolic SolversHow to TrainData-Efficient LLMsHow to Train Data-Efficient LLMsLarge Language Models asOptimizersLarge Language Models as OptimizersThe Llama 3 Herd ofModelsThe Llama 3 Herd of ModelsStudents Rather ThanExperts: A New AI For…Students Rather Than Experts: A New AI For Education Pipeline To Model More Human-Like And Personalised Early AdolescencesA Systematic Assessmentof OpenAI o1-Preview fo…A Systematic Assessment of OpenAI o1-Preview for Higher Order Thinking in EducationReMA: Learning toMeta-Think for LLMs wit…ReMA: Learning to Meta-Think for LLMs with Multi-agent Reinforcement LearningAgentic Large LanguageModels, a SurveyAgentic Large Language Models, a SurveyOn the Power ofContext-Enhanced…On the Power of Context-Enhanced Learning in LLMsMulti-Step Reasoningwith Large Language…Multi-Step Reasoning with Large Language Models, a SurveyR-CoT: ReverseChain-of-Thought Proble…R-CoT: Reverse Chain-of-Thought Problem Generation for Geometric Reasoning in Large Multimodal ModelsAdaptMI: AdaptiveSkill-based In-context…AdaptMI: Adaptive Skill-based In-context Math Instruction for Small Language ModelsUnleashing ReasoningCapability of LLMs via…Unleashing Reasoning Capability of LLMs via Scalable Question Synthesis from ScratchNaturalReasoning:Reasoning in the Wild…NaturalReasoning: Reasoning in the Wild with 2.8M Challenging QuestionsLearnMate: EnhancingOnline Education with…LearnMate: Enhancing Online Education with LLM-Powered Personalized Learning Plans and SupportMetacognitiveCapabilities of LLMs: A…Metacognitive Capabilities of LLMs: An Exploration in Mathematical Problem Solving過去の参考文献中心の論文この論文を引用する論文古い新しい

ノードをクリックするとフォーカスを固定、空白をクリックすると本論文に戻ります。ホバーで一時的にプレビューできます。各ノードのページはタイトルから開けます。