• Home
  • Search
  • A Formal Framework for Evaluating Reasoning Integrity in Language Models
  • https://doi.org/10.20944/preprints202603.2034.v1Copy DOI Icon

A Formal Framework for Evaluating Reasoning Integrity in Language Models

Show More
  • Abstract
  • Literature Map
  • Similar Papers
Abstract

Traditional evaluation of language models prioritizes Ñnal-answer accuracy, offering limited insight into the reasoning processes that produce those outputs. Thispaper introduces a formal framework for evaluating reasoning integrity by modelinginference as a trajectory of belief states under uncertainty. We deÑne externallyobservable belief states that capture hypotheses, uncertainty distributions, and con-straints at each reasoning step, enabling analysis without reliance on internal modelrepresentations. Building on this formulation, we propose a divergence functional that quantiÑessustained disagreement between reasoning trajectories, together with a complexityregularization term that penalizes excessive or redundant reasoning. These compo-nents are combined into a uniÑed scoring function that balances consistency andparsimony. To operationalize the framework, we introduce a multi-stage evalua-tion protocol that constrains intermediate reasoning, injects minimal adversarialperturbations, and measures both divergence and repair cost. We establish theoretical properties of the proposed metrics, including bound-edness, invariance under semantic-preserving transformations, and stability undercontrolled perturbations. Analytical examples illustrate how the framework distin-guishes robust reasoning processes from brittle or superÑcial ones that maintaincorrectness without internal consistency. By shifting evaluation from outcomes tothe dynamics of reasoning, this framework provides a principled basis for assessingreliability and stability in modern language models.

Similar Papers
  • PDF
  • Research Article
  • Citations2

Measuring the Impact of Language Models in Sentiment Analysis for Mexico’s COVID-19 Pandemic

  • Aug 10, 2022
  • Electronics
  • Edgar León-Sandoval +3
  • Preprint Article
  • Citations3

Large Language Models with Novel Token Processing Architecture: A Study of the Dynamic Sequential Transformer

  • Oct 22, 2024
  • Kusi Men +4
  • Conference Article
  • Citations2

Do LLMs learn a true syntactic universal?

  • Jan 01, 2024
  • John T Hale +1
  • Research Article

On Singing Together

  • Oct 14, 2024
  • Philosophy International Journal
  • Allaerts W
  • Research Article

Spiritual-Emotion Guided Fine-Tuning of Mistral-7B for Mental Health Conversational Systems using Bhagavad Gita Knowledge

  • Apr 24, 2026
  • International Scientific Journal of Engineering and Management
  • Sujoy S +3
  • PDF
  • Research Article
  • Citations2

Judgment aggregation, discursive dilemma and reflective equilibrium: Neural language models as self-improving doxastic agents

  • Oct 18, 2022
  • Frontiers in Artificial Intelligence
  • Gregor Betz +1
  • Conference Article
  • Citations5

Active Trajectory Estimation for Partially Observed Markov Decision Processes via Conditional Entropy

  • Jun 29, 2021
  • Timothy L Molloy +1
  • Research Article

Оценка языковой способности нейронных моделей на материале предикативного согласования в русском языке

  • Jan 01, 2022
  • Proceedings of the Institute for System Programming of the RAS
  • Kseniia Andreevna Studenikina
  • Research Article
  • Citations1

ConflLlama: Domain-specific adaptation of large language models for conflict event classification

  • Jul 01, 2025
  • Research & Politics
  • Shreyas Meher +1
  • Conference Article

How Useful is Continued Pre-Training for Generative Unsupervised Domain Adaptation?

  • Jan 01, 2024
  • Rheeya Uppaal +2
  • Book Chapter

Three Issues in Modern Language Modeling

  • Jan 01, 2004
  • Dietrich Klakow
  • Research Article
  • Citations8

The development and validation of a scoring system for shoulder injuries in rugby players

  • Feb 12, 2013
  • British Journal of Sports Medicine
  • Simon Benedict Roberts +1
  • Research Article

Construction of the Mobility to Participation Assessment Scale for Stroke (MPASS) and Testing Its Validity and Reliability in Persons With Stroke in Thailand.

  • Jun 13, 2022
  • Journal of preventive medicine and public health = Yebang Uihakhoe chi
  • Jiraphat Nawarat +1
  • Research Article
  • Citations61

Validation of the Spine Oncology Study Group—Outcomes Questionnaire to assess quality of life in patients with metastatic spine disease

  • Aug 05, 2015
  • The Spine Journal
  • Stein J Janssen +6
  • PDF
  • Research Article
  • Citations7

A Deeper Look at Sheet Music Composer Classification Using Self-Supervised Pretraining

  • Feb 04, 2021
  • Applied Sciences
  • Daniel Yang +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.