Semantic Scene Completion with 2D and 3D Feature Fusion

Sangmin Park; Jong-Eun Ha

doi:10.1109/access.2024.3470754

ScienceGate Book Chapters

JOURNAL ARTICLE

Semantic Scene Completion with 2D and 3D Feature Fusion

Sangmin Park Jong-Eun Ha

Year: 2024 Journal: IEEE Access Pages: 1-1 Publisher: Institute of Electrical and Electronics Engineers

DOI: 10.1109/access.2024.3470754

Get Full-Text PDF Get Analytical Report

Abstract

3D semantic scene completion (SSC) aims to get a dense semantic understanding of an environment in 3D. It requires a geometric and semantic knowledge of the surrounding environment and the filling of void areas. In this paper, we propose an improved algorithm by modifying VoxFormer. VoxFormer consists of two steps for 3D semantic scene completion. First, it predicts the occupancy of an environment. Then, it completes the semantic scene completion through a masked autoencoder. It requires separate training for two stages, which can cause a disconnect of information from input to output. We propose an improved VoxFormer algorithm that makes end-to-end training possible by integrating occupancy prediction and scene completion. We use pseudo-LiDAR computed by depth estimation as input of 3D CNN, which generates queries for cross attention with 2D features. This makes the process end-to-end by connecting occupancy prediction and semantic scene completion. Experimental results using SemanticKITTI show improvement in the proposed algorithm.

Keywords:

Computer science Artificial intelligence Fusion Feature (linguistics) Computer vision Natural language processing Pattern recognition (psychology)

Metrics

Cited By

0.00

FWCI (Field Weighted Citation Impact)

Refs

0.18

Citation Normalized Percentile

Is in top 1%

Is in top 10%

Topics

Video Analysis and Summarization

Physical Sciences → Computer Science → Computer Vision and Pattern Recognition

Multimodal Machine Learning Applications

Physical Sciences → Computer Science → Computer Vision and Pattern Recognition

Advanced Image and Video Retrieval Techniques

Physical Sciences → Computer Science → Computer Vision and Pattern Recognition

Semantic Scene Completion with 2D and 3D Feature Fusion

Abstract

Metrics

Topics

Related Documents

Semantic Scene Completion through Multi-Level Feature Fusion

Semantic Scene Completion with Point Cloud Representation and Transformer-based feature fusion

Multi-Head Multi-Scale Feature Fusion Network for Semantic Scene Completion

AEFF-SSC: An Attention-Enhanced Feature Fusion for 3D Semantic Scene Completion

FFNet: Frequency Fusion Network for Semantic Scene Completion