TASK	DATASET	MODEL	METRIC NAME	METRIC VALUE	GLOBAL RANK
3D Point Cloud Classification	ModelNet40	Point-FEMAE	Overall Accuracy	94.5	# 10
Few-Shot 3D Point Cloud Classification	ModelNet40 10-way (10-shot)	Point-FEMAE	Overall Accuracy	94.0	# 3
Few-Shot 3D Point Cloud Classification	ModelNet40 10-way (20-shot)	Point-FEMAE	Overall Accuracy	95.8	# 3
3D Point Cloud Classification	ScanObjectNN	Point-FEMAE	Overall Accuracy	90.22	# 13
3D Point Cloud Classification	ScanObjectNN	Point-FEMAE	OBJ-BG (OA)	95.18	# 5
3D Point Cloud Classification	ScanObjectNN	Point-FEMAE	OBJ-ONLY (OA)	93.29	# 5

Badge	Markdown
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/towards-compact-3d-representations-via-point/few-shot-3d-point-cloud-classification-on-3)](https://paperswithcode.com/sota/few-shot-3d-point-cloud-classification-on-3?p=towards-compact-3d-representations-via-point)`
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/towards-compact-3d-representations-via-point/few-shot-3d-point-cloud-classification-on-4)](https://paperswithcode.com/sota/few-shot-3d-point-cloud-classification-on-4?p=towards-compact-3d-representations-via-point)`
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/towards-compact-3d-representations-via-point/3d-point-cloud-classification-on-modelnet40)](https://paperswithcode.com/sota/3d-point-cloud-classification-on-modelnet40?p=towards-compact-3d-representations-via-point)`
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/towards-compact-3d-representations-via-point/3d-point-cloud-classification-on-scanobjectnn)](https://paperswithcode.com/sota/3d-point-cloud-classification-on-scanobjectnn?p=towards-compact-3d-representations-via-point)`

Towards Compact 3D Representations via Point Feature Enhancement Masked Autoencoders

17 Dec 2023 · Yaohua Zha, Huizhen Ji, Jinmin Li, Rongsheng Li, Tao Dai, Bin Chen, Zhi Wang, Shu-Tao Xia ·

Learning 3D representation plays a critical role in masked autoencoder (MAE) based pre-training methods for point cloud, including single-modal and cross-modal based MAE. Specifically, although cross-modal MAE methods learn strong 3D representations via the auxiliary of other modal knowledge, they often suffer from heavy computational burdens and heavily rely on massive cross-modal data pairs that are often unavailable, which hinders their applications in practice. Instead, single-modal methods with solely point clouds as input are preferred in real applications due to their simplicity and efficiency. However, such methods easily suffer from limited 3D representations with global random mask input. To learn compact 3D representations, we propose a simple yet effective Point Feature Enhancement Masked Autoencoders (Point-FEMAE), which mainly consists of a global branch and a local branch to capture latent semantic features. Specifically, to learn more compact features, a share-parameter Transformer encoder is introduced to extract point features from the global and local unmasked patches obtained by global random and local block mask strategies, followed by a specific decoder to reconstruct. Meanwhile, to further enhance features in the local branch, we propose a Local Enhancement Module with local patch convolution to perceive fine-grained local context at larger scales. Our method significantly improves the pre-training efficiency compared to cross-modal alternatives, and extensive downstream experiments underscore the state-of-the-art effectiveness, particularly outperforming our baseline (Point-MAE) by 5.16%, 5.00%, and 5.04% in three variants of ScanObjectNN, respectively. The code is available at https://github.com/zyh16143998882/AAAI24-PointFEMAE.

PDF Abstract

Code

Add Remove Mark official

zyh16143998882/aaai24-pointfemae official

Tasks

Add Remove

3D Point Cloud Classification

Few-Shot 3D Point Cloud Classification

Datasets

ShapeNet

ModelNet

ScanObjectNN

Results from the Paper

Add Remove

Ranked #3 on Few-Shot 3D Point Cloud Classification on ModelNet40 10-way (20-shot) (using extra training data)

Get a GitHub badge

Task	Dataset	Model	Metric Name	Metric Value	Global Rank	Benchmark
3D Point Cloud Classification	ModelNet40	Point-FEMAE	Overall Accuracy	94.5	# 10	Compare
Few-Shot 3D Point Cloud Classification	ModelNet40 10-way (10-shot)	Point-FEMAE	Overall Accuracy	94.0	# 3	Compare
Few-Shot 3D Point Cloud Classification	ModelNet40 10-way (20-shot)	Point-FEMAE	Overall Accuracy	95.8	# 3	Compare
3D Point Cloud Classification	ScanObjectNN	Point-FEMAE	Overall Accuracy	90.22	# 13	Compare
			OBJ-BG (OA)	95.18	# 5	Compare
			OBJ-ONLY (OA)	93.29	# 5	Compare

Methods

Add Remove

Absolute Position Encodings • Adam • AutoEncoder • BPE • Convolution • Dense Connections • Dropout • Label Smoothing • Layer Normalization • Linear Layer • MAE • Multi-Head Attention • Position-Wise Feed-Forward Layer • Residual Connection • Scaled Dot-Product Attention • Softmax • Transformer

Edit Social Preview

Towards Compact 3D Representations via Point Feature Enhancement Masked Autoencoders

Code Edit Add Remove Mark official

Tasks Edit Add Remove

Datasets Edit

Results from the Paper Edit Add Remove

Methods Edit Add Remove

Code

Add Remove Mark official

Tasks

Add Remove

Datasets

Results from the Paper

Add Remove

Methods

Add Remove