• Home
  • Search
  • Mathematical discoveries from program search with large language models
  • Cite Icon305
  • https://doi.org/10.1038/s41586-023-06924-6Copy DOI Icon

Mathematical discoveries from program search with large language models

  • Dec 14, 2023
  • Nature
  • Bernardino Romera-Paredes +11 more
Show More
  • Abstract
  • Highlights & Summary
  • PDF
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Large language models (LLMs) have demonstrated tremendous capabilities in solving complex tasks, from quantitative reasoning to understanding natural language. However, LLMs sometimes suffer from confabulations (or hallucinations), which can result in them making plausible but incorrect statements1,2. This hinders the use of current large models in scientific discovery. Here we introduce FunSearch (short for searching in the function space), an evolutionary procedure based on pairing a pretrained LLM with a systematic evaluator. We demonstrate the effectiveness of this approach to surpass the best-known results in important problems, pushing the boundary of existing LLM-based approaches3. Applying FunSearch to a central problem in extremal combinatorics—the cap set problem—we discover new constructions of large cap sets going beyond the best-known ones, both in finite dimensional and asymptotic cases. This shows that it is possible to make discoveries for established open problems using LLMs. We showcase the generality of FunSearch by applying it to an algorithmic problem, online bin packing, finding new heuristics that improve on widely used baselines. In contrast to most computer search approaches, FunSearch searches for programs that describe how to solve a problem, rather than what the solution is. Beyond being an effective and scalable strategy, discovered programs tend to be more interpretable than raw solutions, enabling feedback loops between domain experts and FunSearch, and the deployment of such programs in real-world applications.

Loading PDF

Similar Papers
  • Research Article

TencentLLMEval: A Hierarchical Evaluation of Real-World Capabilities for Human-Aligned LLMs

  • Apr 29, 2025
  • ACM Transactions on Intelligent Systems and Technology
  • Shuyi Xie +15
  • Supplementary Content
  • Citations114

Applications and Concerns of ChatGPT and Other Conversational Large Language Models in Health Care: Systematic Review

  • Nov 07, 2024
  • Journal of Medical Internet Research
  • Leyao Wang +7
  • Research Article

Methodology for Creating a Benchmark to Evaluate LLM Performance on Numerals

  • Dec 04, 2025
  • Информатика и автоматизация
  • Sergey Karpovich +2
  • Preprint Article
  • Citations1

Utilizing Large Language Models for Geoscience Literature Information Extraction

  • Jan 20, 2025
  • Peng Yu +3
  • Conference Article

Community Member and Scientific Queries for Climate Resilience: An Assessment of the LLM Landscape

  • Nov 12, 2025
  • Rhoda Nankabirwa +1
  • Research Article
  • Citations1

Large Language Model Powered Symbolic Execution

  • Oct 09, 2025
  • Proceedings of the ACM on Programming Languages
  • Yihe Li +2
  • Conference Article
  • Citations4

NegativePrompt: Leveraging Psychology for Large Language Models Enhancement via Negative Emotional Stimuli

  • Aug 01, 2024
  • Xu Wang +2
  • Conference Article
  • Citations1

Intuitive or Dependent? Investigating LLMs’ Behavior Style to Conflicting Prompts

  • Jan 01, 2024
  • Proceedings of the Annual Meeting of the Association for Computational Linguistics
  • Jiahao Ying +5
  • Research Article

Look Before You Leap: Enhance Attention and Vigilance Regarding Harmful Content with GuidelineLLM

  • Apr 11, 2025
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Shaoqing Zhang +6
  • Research Article
  • Citations4

Leveraging LLMs for Automated Extraction and Structuring of Educational Concepts and Relationships

  • Sep 19, 2025
  • Machine Learning and Knowledge Extraction
  • Tianyuan Yang +5
  • Research Article
  • Citations2

Toward Cross-Hospital Deployment of Natural Language Processing Systems: Model Development and Validation of Fine-Tuned Large Language Models for Disease Name Recognition in Japanese

  • Jul 08, 2025
  • JMIR Medical Informatics
  • Seiji Shimizu +4
  • PDF
  • Research Article
  • Citations15

Cost-efficient prompt engineering for unsupervised entity resolution in the product matching domain

  • Aug 16, 2024
  • Discover Artificial Intelligence
  • Navapat Nananukul +2
  • Conference Article
  • Citations3

Graph Reasoning with LLMs (GReaL)

  • Aug 24, 2024
  • Anton Tsitsulin +3
  • Conference Article
  • Citations57

Extract, Define, Canonicalize: An LLM-based Framework for Knowledge Graph Construction

  • Jan 01, 2024
  • Bowen Zhang +1
  • Research Article

Large language models in healthcare and biomedical informatics: A comprehensive review

  • Jan 01, 2026
  • Innovation and Emerging Technologies
  • Andrew Hornback +8
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.