• Home
  • Search
  • A Survey on Large Language Models for Mathematical Reasoning
  • https://doi.org/10.1145/3786333Copy DOI Icon

A Survey on Large Language Models for Mathematical Reasoning

Show More
  • Abstract
  • Literature Map
  • References
  • Similar Papers
Abstract

Mathematical reasoning has long represented one of the most fundamental and challenging frontiers in artificial intelligence research. In recent years, large language models (LLMs) have achieved significant advances in this area. This survey examines the development of mathematical reasoning abilities in LLMs through two high-level cognitive phases: comprehension, where models gain mathematical understanding via diverse pretraining strategies, and answer generation, which has progressed from direct prediction to step-by-step Chain-of-Thought (CoT) reasoning. We review methods for enhancing mathematical reasoning, ranging from training-free prompting to fine-tuning approaches such as supervised fine-tuning and reinforcement learning, and discuss recent work on extended CoT and “test-time scaling”. Despite notable progress, fundamental challenges remain in terms of capacity, efficiency, and generalization. To address these issues, we highlight promising research directions, including advanced pretraining and knowledge augmentation techniques, formal reasoning frameworks, and meta-generalization through principled learning paradigms. This survey tries to provide some insights for researchers interested in enhancing reasoning capabilities of LLMs and for those seeking to apply these techniques to other domains.

Similar Papers
  • Research Article
  • Citations3

TWOSOME: An Efficient Online Framework to Align LLMs with Embodied Environments via Reinforcement Learning

  • Jun 01, 2024
  • International Journal of Artificial Intelligence and Robotics Research
  • Weihao Tan +5
  • Research Article
  • Citations1

Assessing the Creativity of LLMs in Proposing Novel Solutions to Mathematical Problems

  • Apr 11, 2025
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Junyi Ye +4
  • Research Article
  • Citations1

RMath: A Logic Reasoning-Focused Datasets Toward Mathematical Multistep Reasoning Tasks

  • Apr 11, 2025
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Ziyi Hu +5
  • PDF
  • Research Article
  • Citations6

MathOdyssey: Benchmarking Mathematical Problem-Solving Skills in Large Language Models Using Odyssey Math Data

  • Aug 08, 2025
  • Scientific Data
  • Meng Fang +4
  • Research Article

Look Before You Leap: Enhance Attention and Vigilance Regarding Harmful Content with GuidelineLLM

  • Apr 11, 2025
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Shaoqing Zhang +6
  • Research Article
  • Citations56

Generative AI in cybersecurity: A comprehensive review of LLM applications and vulnerabilities

  • Jan 01, 2025
  • Internet of Things and Cyber-Physical Systems
  • Mohamed Amine Ferrag +7
  • Research Article
  • Citations51

Large language models for biomedicine: foundations, opportunities, challenges, and best practices.

  • Apr 24, 2024
  • Journal of the American Medical Informatics Association : JAMIA
  • Satya S Sahoo +8
  • Research Article
  • Citations1

Learning Theorem Rationale for Improving the Mathematical Reasoning Capability of Large Language Models

  • Apr 11, 2025
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Yu Sheng +2
  • Research Article
  • Citations4

Beyond Text Generation: Assessing Large Language Models’ Ability to Reason Logically and Follow Strict Rules

  • Jan 15, 2025
  • AI
  • Zhiyong Han +4
  • Research Article
  • Citations281

The debate over understanding in AI’s large language models

  • Mar 21, 2023
  • Proceedings of the National Academy of Sciences of the United States of America
  • Melanie Mitchell +1
  • Conference Article
  • Citations8

RLHFPoison: Reward Poisoning Attack for Reinforcement Learning with Human Feedback in Large Language Models

  • Jan 01, 2024
  • Jiongxiao Wang +4
  • Supplementary Content
  • Citations26

Toward large reasoning models: A survey of reinforced reasoning with large language models

  • Oct 01, 2025
  • Patterns
  • Fengli Xu +19
  • Conference Article
  • Citations4

Step-level Value Preference Optimization for Mathematical Reasoning

  • Jan 01, 2024
  • Guoxin Chen +3
  • Research Article

AI-Powered Automated Penetration Testing in Kali Linux: An Enterprise-Scale Offensive Security Framework Driven by Reinforcement Learning and Large Language Models

  • Feb 05, 2026
  • International Journal of Science and Research (IJSR)
  • Harunmiya S Malek
  • Research Article
  • Citations1

Diversity of Thought Elicits Stronger Reasoning Capabilities in Multi-Agent Debate Frameworks

  • Dec 13, 2024
  • Journal of Robotics and Automation Research
  • Mahmood Hegazy
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.