Journal Article10.1109/cvprw59228.2023.00025
TorchSparse++: Efficient Point Cloud Engine
Haotian Tang,Shang Yang,Zhijian Liu,Ke Hong,Zhongming Yu,Xiuyu Li,Guohao Dai,Yu Wang,Song Han +8 more
- 01 Jun 2023
pp 202-209
10
TL;DR: This work systematically analyze and improve existing dataflows for convolution on point clouds and achieves end-to-end speedup on an NVIDIA A100 GPU over the state-of-the-art MinkowskiEngine, SpConv 1.2, TorchSparse and SpCon v2 in inference respectively.
read more
Abstract: Point cloud computation has become an increasingly more important workload for autonomous driving and other applications. Unlike dense 2D computation, point cloud convolution has sparse and irregular computation patterns and thus requires dedicated inference system support with specialized high-performance kernels. While existing point cloud deep learning libraries have developed different dataflows for convolution on point clouds, they assume a single dataflow throughout the execution of the entire model. In this work, we systematically analyze and improve existing dataflows. Our resulting system, TorchSparse++, achieves 2.9×, 3.3×, 2.2× and 1.8× measured end-to-end speedup on an NVIDIA A100 GPU over the state-of-the-art MinkowskiEngine, SpConv 1.2, TorchSparse and SpConv v2 in inference respectively. Furthermore, TorchSparse++ is the only system to date that supports all necessary primitives for 3D segmentation, detection, and reconstruction workloads in autonomous driving. Code is publicly released at https://github.com/mit-han-lab/torchsparse.
read more
Chat with Paper
AI Agents for this Paper
Find similar papers on Google Scholar, PubMed and Arxiv
Write a critical review of this paper
Analyze citations of this paper to find unaddressed research gaps
Citations
Robo3D: Towards Robust and Reliable 3D Perception against Corruptions
Ling‐Dong Kong,Youquan Liu,Xin Li,Runnan Chen,Wenwei Zhang,Jianxin Ren,Liang Peng,Kai Chen,Ziwei Liu +8 more
- 01 Oct 2023
TL;DR: Robo3D is a benchmark for probing the robustness of 3D detectors and segmentors against natural corruptions. It includes eight corruption types and a density-insensitive training framework to enhance model resilience.
28
PointConvFormer: Revenge of the Point-based Convolution
Wenxuan Wu,Fuxin Li,Shihua Qi +2 more
- 01 Jun 2023
TL;DR: PointConvFormer is a novel point cloud based deep network architecture that combines point convolution and Transformers. It achieves better accuracy-speed tradeoff than classic convolutions, regular transformers, and voxelized sparse convolution approaches.
21
SparseViT: Revisiting Activation Sparsity for Efficient High-Resolution Vision Transformer
Xuanyao Chen,Zhijian Liu,Haotian Tang,Yi Li,Hang Zhao,Song Han +5 more
- 01 Jun 2023
TL;DR: SparseViT revisits activation sparsity for high-resolution vision transformers, achieving significant latency reduction with minimal accuracy loss.
18
TUMTraf V2X Cooperative Perception Dataset
Walter Zimmer,Gerhard Arya Wardana,Suren Sritharan,Xingcheng Zhou,Rui Song,Alois Knoll +5 more
- 16 Jun 2024
8
Learning Human Mesh Recovery in 3D Scenes
Zehong Shen,Zhi Cen,Sida Peng,Qing Shuai,Hujun Bao,Xiaowei Zhou +5 more
- 01 Jun 2023
TL;DR: A novel method for recovering human mesh in 3D scenes from a single image. Estimates absolute position and dense scene contacts with a sparse 3D CNN, and enhances a pretrained human mesh recovery network by cross-attention with the derived 3D scene cues.
7
References
•Proceedings Article
PointNet++: Deep Hierarchical Feature Learning on Point Sets in a Metric Space
Charles R. Qi,Li Yi,Hao Su,Leonidas J. Guibas +3 more
- 07 Jun 2017
TL;DR: PointNet++ as discussed by the authors applies PointNet recursively on a nested partitioning of the input point set to learn local features with increasing contextual scales, and proposes novel set learning layers to adaptively combine features from multiple scales.
nuScenes: A Multimodal Dataset for Autonomous Driving
Holger Caesar,Varun Bankiti,Alex H. Lang,Sourabh Vora,Venice Erin Liong,Qiang Xu,Anush Krishnan,Yu Pan,Giancarlo Baldan,Oscar Beijbom +9 more
- 14 Jun 2020
TL;DR: nuScenes as discussed by the authors is the first dataset to carry the full autonomous vehicle sensor suite: 6 cameras, 5 radars and 1 lidar, all with full 360 degree field of view.
•Posted Content
nuScenes: A multimodal dataset for autonomous driving
Holger Caesar,Varun Bankiti,Alex H. Lang,Sourabh Vora,Venice Erin Liong,Qiang Xu,Anush Krishnan,Yu Pan,Giancarlo Baldan,Oscar Beijbom +9 more
TL;DR: nuScenes as mentioned in this paper is the first dataset to carry the full autonomous vehicle sensor suite: 6 cameras, 5 radars and 1 lidar, all with full 360 degree field of view.
3.7K
SECOND: Sparsely Embedded Convolutional Detection
Yan Yan,Yuxing Mao,Bo Li +2 more
TL;DR: An improved sparse convolution method for Voxel-based 3D convolutional networks is investigated, which significantly increases the speed of both training and inference and introduces a new form of angle loss regression to improve the orientation estimation performance.
3.2K
PointPillars: Fast Encoders for Object Detection From Point Clouds
Alex H. Lang,Sourabh Vora,Holger Caesar,Lubing Zhou,Jiong Yang,Oscar Beijbom +5 more
- 15 Jun 2019
TL;DR: benchmarks suggest that PointPillars is an appropriate encoding for object detection in point clouds, and proposes a lean downstream network.