• Home
  • Search
  • OctFormer: Octree-based Transformers for 3D Point Clouds
  • Cite Icon85
  • https://doi.org/10.1145/3592131Copy DOI Icon

OctFormer: Octree-based Transformers for 3D Point Clouds

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

We propose octree-based transformers, named OctFormer, for 3D point cloud learning. OctFormer can not only serve as a general and effective backbone for 3D point cloud segmentation and object detection but also have linear complexity and is scalable for large-scale point clouds. The key challenge in applying transformers to point clouds is reducing the quadratic, thus overwhelming, computation complexity of attentions. To combat this issue, several works divide point clouds into non-overlapping windows and constrain attentions in each local window. However, the point number in each window varies greatly, impeding the efficient execution on GPU. Observing that attentions are robust to the shapes of local windows, we propose a novel octree attention, which leverages sorted shuffled keys of octrees to partition point clouds into local windows containing a fixed number of points while permitting shapes of windows to change freely. And we also introduce dilated octree attention to expand the receptive field further. Our octree attention can be implemented in 10 lines of code with open-sourced libraries and runs 17 times faster than other point cloud attentions when the point number exceeds 200 k. Built upon the octree attention, OctFormer can be easily scaled up and achieves state-of-the-art performances on a series of 3D semantic segmentation and 3D object detection benchmarks, surpassing previous sparse-voxel-based CNNs and point cloud transformers in terms of both efficiency and effectiveness. Notably, on the challenging ScanNet200 dataset, OctFormer outperforms sparse-voxel-based CNNs by 7.3 in mIoU. Our code and trained models are available at https://wang-ps.github.io/octformer.

Similar Papers
  • Conference Article
  • Citations104

Image-to-Lidar Self-Supervised Distillation for Autonomous Driving Data

  • Jun 01, 2022
  • Corentin Sautier +5
  • PDF
  • Research Article
  • Citations4

Enhancing Point Features with Spatial Information for Point-Based 3D Object Detection

  • Dec 21, 2021
  • Scientific Programming
  • Huaijin Liu +3
  • Conference Article
  • Citations52

FVNet: 3D Front-View Proposal Generation for Real-Time Object Detection from Point Clouds

  • Oct 01, 2019
  • Jie Zhou +3
  • Conference Article
  • Citations14

3DYOLO: Real-time 3D Object Detection in 3D Point Clouds for Autonomous Driving

  • Dec 06, 2021
  • Priya M V +1
  • PDF
  • Research Article
  • Citations3

Density Awareness and Neighborhood Attention for LiDAR-Based 3D Object Detection

  • Nov 02, 2022
  • Photonics
  • Hanxiang Qian +4
  • Conference Article
  • Citations3

U-Select RCNN: An Effective Voxel-based 3D Object Detection Method with Feature Selection Strategy

  • Aug 15, 2022
  • Zhenghong Zhang +3
  • Research Article
  • Citations8

PLOT: a 3D point cloud object detection network for autonomous driving

  • Jan 16, 2023
  • Robotica
  • Yihuan Zhang +2
  • Research Article
  • Citations43

RI-Fusion: 3D Object Detection Using Enhanced Point Features With Range-Image Fusion for Autonomous Driving

  • Jan 01, 2023
  • IEEE Transactions on Instrumentation and Measurement
  • Xinyu Zhang +8
  • Research Article
  • Citations2

3D object detection combining semantic and geometric features from point clouds

  • Jan 12, 2022
  • Cobot
  • Hao Peng +4
  • Research Article

An overview of real-time 3D object detection for mobile robots: towards multi-sensor fusion

  • Sep 12, 2025
  • International Journal of Systems Science
  • Bowen Chen +4
  • PDF
  • Research Article
  • Citations9

Pillar-Based 3D Object Detection from Point Cloud with Multiattention Mechanism

  • Feb 09, 2023
  • Wireless Communications and Mobile Computing
  • Xin Li +7
  • Research Article
  • Citations36

SP-Det: Leveraging Saliency Prediction for Voxel-Based 3D Object Detection in Sparse Point Cloud

  • Jan 01, 2024
  • IEEE Transactions on Multimedia
  • Pei An +7
  • Conference Article
  • Citations47

Masked Autoencoder for Self-Supervised Pre-training on Lidar Point Clouds

  • Jan 01, 2023
  • Georg Hess +5
  • Conference Article
  • Citations1

SparseDet: Towards End-to-End 3D Object Detection

  • Jan 01, 2022
  • Jianhong Han +4
  • Research Article
  • Citations2

Data Augmentation in 3D Object Detection for self-driving vehicles: the role of original and augmented training samples

  • Dec 01, 2022
  • Journal of Physics: Conference Series
  • X Santos +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.