• Home
  • Search
  • M3ENet: A Multi-Modal Fusion Network for Efficient Micro-Expression Recognition
  • Cite Icon2
  • https://doi.org/10.3390/s25206276Copy DOI Icon

M3ENet: A Multi-Modal Fusion Network for Efficient Micro-Expression Recognition

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Micro-expression recognition (MER) aims to detect brief and subtle facial movements that reveal suppressed emotions, discerning authentic emotional responses in scenarios such as visitor experience analysis in museum settings. However, it remains a highly challenging task due to the fleeting duration, low intensity, and limited availability of annotated data. Most existing approaches rely solely on either appearance or motion cues, thereby restricting their ability to capture expressive information fully. To overcome these limitations, we propose a lightweight multi-modal fusion network, termed M3ENet, which integrates both motion and appearance cues through early-stage feature fusion. Specifically, our model extracts horizontal, vertical, and strain-based optical flow between the onset and apex frames, alongside RGB images from the onset, apex, and offset frames. These inputs are processed by two modality-specific subnetworks, whose features are fused to exploit complementary information for robust classification. To improve generalization in low data regimes, we employ targeted data augmentation and adopt focal loss to mitigate class imbalance. Extensive experiments on five benchmark datasets, including CASME I, CASME II, CAS(ME)2, SAMM, and MMEW, demonstrate that M3ENet achieves state-of-the-art performance with high efficiency. Ablation studies and Grad-CAM visualizations further confirm the effectiveness and interpretability of the proposed architecture.

Similar Papers
  • Research Article

Multimodal fusion network with learnable wavelet-enhanced features for hyperspectral unmixing

  • Oct 30, 2025
  • International Journal of Remote Sensing
  • Zhixiang Wang +4
  • PDF
  • Research Article
  • Citations17

Micro-Expression Recognition Based on Optical Flow and PCANet+

  • Jun 05, 2022
  • Sensors (Basel, Switzerland)
  • Shiqi Wang +5
  • Research Article
  • Citations399

Less is more: Micro-expression recognition from video using apex frame

  • Dec 14, 2017
  • Signal Processing: Image Communication
  • Sze-Teng Liong +3
  • Book Chapter
  • Citations5

Shot-Based Hybrid Fusion for Movie Genre Classification

  • Jan 01, 2022
  • Tianyu Bi +2
  • Research Article
  • Citations433

Towards Reading Hidden Emotions: A Comparative Study of Spontaneous Micro-Expression Spotting and Recognition Methods

  • Oct 01, 2018
  • IEEE Transactions on Affective Computing
  • Xiaobai Li +6
  • Research Article
  • Citations6

A multimodal dual-branch fusion network for fetal hypoxia detection

  • Aug 29, 2024
  • Expert Systems With Applications
  • Mujun Liu +5
  • Research Article
  • Citations31

MMHFNet: Multi-modal and multi-layer hybrid fusion network for voice pathology detection

  • Mar 14, 2023
  • Expert Systems with Applications
  • Hussein M.A Mohammed +2
  • Research Article
  • Citations22

Multimodal Co-Attention Fusion Network With Online Data Augmentation for Cancer Subtype Classification.

  • Nov 01, 2024
  • IEEE transactions on medical imaging
  • Saisai Ding +4
  • Research Article
  • Citations1

MMFNet: A Mamba-Based Multimodal Fusion Network for Remote Sensing Image Semantic Segmentation.

  • Oct 08, 2025
  • Sensors (Basel, Switzerland)
  • Jingting Qiu +4
  • Research Article
  • Citations1

Multi-Modal Fusion Network with Multi-Head Self-Attention for Injection Training Evaluation in Medical Education

  • Sep 30, 2024
  • Electronics
  • Zhe Li +4
  • Research Article

A 1D Snoring Waveform and 2D Composite Acoustic Feature Graph-Based Multi-Modal Fusion Network for Obstructive Sites Recognition

  • Jan 01, 2026
  • IEEE Journal of Biomedical and Health Informatics
  • Xia Hu +5
  • PDF
  • Research Article

Remb: regularized embedding memory book to extend metric learning in fault diagnosis

  • Oct 11, 2025
  • Journal of Intelligent Manufacturing
  • David Solís-Martín +3
  • Conference Article
  • Citations8

Go with the Flow: Improving Multi-view Vehicle Detection with Motion Cues

  • Aug 01, 2014
  • Alfredo Ramirez +2
  • Research Article

Micro-Expression Recognition via LoRA-Enhanced DinoV2 and Interactive Spatio-Temporal Modeling

  • Jan 16, 2026
  • Sensors (Basel, Switzerland)
  • Meng Wang +3
  • Research Article
  • Citations89

Deep learning based solar radiation micro forecast by fusion of infrared cloud images and radiation data

  • Apr 28, 2021
  • Applied Energy
  • Meenu Ajith +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.