• Home
  • Search
  • MeTHanol: Modularized Thinking Language Models with Intermediate Layer Thinking, Decoding and Bootstrapping Reasoning
  • https://doi.org/10.1109/ijcnn64981.2025.11229297Copy DOI Icon

MeTHanol: Modularized Thinking Language Models with Intermediate Layer Thinking, Decoding and Bootstrapping Reasoning

  • Jun 30, 2025
  • Ningyuan Xi +5 more
Show More
  • Abstract
  • Literature Map
  • References
  • Similar Papers
Abstract

Current research efforts are focused on enhancing the thinking and reasoning capability of large language model (LLM) by prompting, data-driven emergence and inference-time computation. In this study, we consider stimulating language model’s thinking and cognitive abilities from a modular perspective, which mimics the human brain architecture. We select a specific intermediate attention layer with newly implemented language heads. We conduct dual-layer fine-tuning by annotated (query, thought, answer) samples and show that the intermediate layer can also learn to decode fluent and reasonable language tokens. A two-pass inference mechanism is designed to generate thoughts then formal responses. The entire framework is called modularized thinking language model (MeTHanol) which can enhance LLM’s cognitive behaviors as indicated by Theory of Mind (ToM) and Vignette-based experiments. Case studies also show that MeTHanol can plan and self-reflect and generate human-like thoughts and answers, even on unseen and open-domain tasks. MeTHanol can also adapt to a personalized prompt and behave as the specified character. Our study holds promise for significant cognitive gains from a modular perspective. Our code, model and data are available at https://bachozean.github.io/methanol-page/.

Similar Papers
  • PDF
  • Research Article
  • Citations13

Multi-Level Representation Learning for Chinese Medical Entity Recognition: Model Development and Validation.

  • May 04, 2020
  • JMIR Medical Informatics
  • Zhichang Zhang +2
  • PDF
  • Conference Article
  • Citations63

Contrastive Distillation on Intermediate Representations for Language Model Compression

  • Jan 01, 2020
  • Siqi Sun +5
  • Supplementary Content

Proceedings of the Twelfth Workshop on Asian Translation (WAT 2025)

  • Jan 01, 2025
  • Conference Article

Not All Layers of LLMs Are Necessary During Inference

  • Sep 01, 2025
  • Siqi Fan +7
  • Conference Article
  • Citations2

Towards Robust & Reusable Evaluation for Novelty & Diversity

  • Nov 03, 2014
  • Kai Hui
  • Conference Article

Sorted LLaMA: Unlocking the Potential of Intermediate Layers of Large Language Models for Dynamic Inference

  • Jan 01, 2024
  • Parsa Kavehzadeh +5
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.