• Home
  • Search
  • High Performance Hierarchical Tucker Tensor Learning Using GPU Tensor Cores
  • Cite Icon16
  • https://doi.org/10.1109/tc.2022.3172895Copy DOI Icon

High Performance Hierarchical Tucker Tensor Learning Using GPU Tensor Cores

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Extracting information from large-scale high-dimensional data is a fundamentally important task in high performance computing, where the hierarchical Tucker (HT) tensor learning approach (learning a tensor-tree structure) has been widely used in many applications. However, HT tensor learning algorithms are compute-intensive due to the " <i>curse of dimensionality</i> ," i.e., the time complexity grows exponentially with the order of the data tensor. The computation of HT tensor learning algorithms boils down to tensor primitives, which are amenable to computing on GPU tensor cores. Existing work does not support HT tensor learning using GPU tensor cores. There are three main challenges to address: 1) to accelerate tensor learning primitives using GPU tensor cores; 2) to implement the tensor learning algorithms using GPU tensor cores and multiple GPUs; 3) to support large-scale data tensors exceeding the GPU memory capacity. In this paper, we present efficient HT tensor learning primitives using GPU tensor cores and demonstrate three applications. First, we utilize GPU tensor cores to optimize HT tensor learning primitives, including tensor contractions, tensor matricizations and tensor singular value decomposition (SVD). We employ the optimized primitives to optimize HT tensor decomposition algorithms for Big Data analysis. Second, we propose a novel HT tensor layer for deep neural networks, whose training process only involves a forward pass without back propagation. The forward pass consists of tensor operations, thus further exploiting the computing power of GPU tensor cores. Third, we apply the optimized primitives to develop a tensor-tree structured quantum machine learning algorithm <i>tree-tensor network (TTN)</i> . Compared with TensorLy and TensorNetwork on NVIDIA A100 GPUs, our third-order HT tensor decomposition algorithm achieves up to <inline-formula><tex-math notation="LaTeX">$8.92 \times$</tex-math></inline-formula> and <inline-formula><tex-math notation="LaTeX">$6.42 \times$</tex-math></inline-formula> speedups, respectively, and our high-order case achieves up to <inline-formula><tex-math notation="LaTeX">$32.67 \times$</tex-math></inline-formula> and <inline-formula><tex-math notation="LaTeX">$23.97 \times$</tex-math></inline-formula> speedups, respectively. Our HT tensor layer for a fully connected neural network achieves <inline-formula><tex-math notation="LaTeX">$49.2 \times$</tex-math></inline-formula> compression at the cost of 0.5% drops in accuracy and <inline-formula><tex-math notation="LaTeX">$1.42 \times$</tex-math></inline-formula> speedup compared with the implementation on CUDA cores; for the AlexNet, our HT tensor layer achieves <inline-formula><tex-math notation="LaTeX">$9.45 \times$</tex-math></inline-formula> compression at the cost of 0.8% drops in accuracy and <inline-formula><tex-math notation="LaTeX">$1.87 \times$</tex-math></inline-formula> speedup compared with the implementation on CUDA cores. Our TTN algorithm achieves up to <inline-formula><tex-math notation="LaTeX">$11.17\times$</tex-math></inline-formula> speedup compared with TensorNetwork, indicating the potential of optimized tensor learning primitives for the classical simulation of quantum machine learning algorithms.

Similar Papers
  • Research Article
  • Citations63

Nonconvex low-rank tensor approximation with graph and consistent regularizations for multi-view subspace learning

  • Feb 14, 2023
  • Neural Networks
  • Baicheng Pan +2
  • Book Chapter

Practical Implications of Dequantization on Machine Learning Algorithms: A Survey

  • Jan 01, 2023
  • Vinooth Rao Kulkarni +4
  • Research Article

Define, refined and re-defined concepts of quantum machine learning : A review

  • Jan 01, 2024
  • Journal of Information and Optimization Sciences
  • Shradha Chavan +1
  • Research Article
  • Citations108

Tensor-tensor algebra for optimal representation and compression of multiway data

  • Jul 07, 2021
  • Proceedings of the National Academy of Sciences
  • Misha E Kilmer +3
  • Book Chapter

Quantum Machine Learning Architecture for EEG-Based Emotion Recognition

  • Sep 27, 2024
  • C U Om Kumar +5
  • Research Article

Towards network subspace clustering using low-rank tensor singular value decomposition and graph-boosting

  • Mar 11, 2026
  • Journal of Big Data
  • Jinquan Zhang +1
  • Research Article
  • Citations4

Improved robust tensor principal component analysis for accelerating dynamic MR imaging reconstruction.

  • May 05, 2020
  • Medical &amp; Biological Engineering &amp; Computing
  • Mingfeng Jiang +6
  • PDF
  • Research Article
  • Citations11

Implementation and Performance Evaluation of Quantum Machine Learning Algorithms for Binary Classification

  • Nov 28, 2024
  • Software
  • Surajudeen Shina Ajibosin +1
  • Book Chapter
  • Citations5

Quanvolutional Neural Network Applied to MNIST

  • Jan 01, 2023
  • Daniel Alejandro Lopez +4
  • Research Article
  • Citations1

An orthogonal equivalence theorem for third order tensors

  • Jan 01, 2022
  • Journal of Industrial and Management Optimization
  • Liqun Qi +3
  • Research Article

Online learning algorithms : For passivity-based and distributed control

  • May 03, 2016
  • Research Repository (Delft University of Technology)
  • Subramanya Nageshrao
  • Research Article
  • Citations1

Learning verb complements for Modern Greek: balancing the noisy dataset

  • Jan 01, 2008
  • Natural Language Engineering
  • Katia Kermanidis +3
  • Research Article
  • Citations59

A Tensor-Train Deep Computation Model for Industry Informatics Big Data Feature Learning

  • Jul 01, 2018
  • IEEE Transactions on Industrial Informatics
  • Qingchen Zhang +3
  • Research Article

Investigating robustness dynamics of shallow machine learning techniques: beyond basic natural variations in image classification

  • Dec 26, 2024
  • Connection Science
  • Mengtong Li
  • Book Chapter
  • Citations7

Bridging Few-Shot Learning and Adaptation: New Challenges of Support-Query Shift

  • Jan 01, 2021
  • Etienne Bennequin +4
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.