TensorX
返回文献探索

Paper · arXiv 2412.03515

Distilling Diffusion Models to Efficient 3D LiDAR Scene Completion

Shengyuan Zhang, An Zhao, Ling Yang, Zejian Li, Chenye Meng, Haoran Xu, Tianrun Chen, AnYang Wei, Perry Pengyun GU, Lingyun Sun

27 upvotesDecember 4, 2024arXiv 预印本
AI 摘要

A novel distillation method, ScoreLiDAR, accelerates 3D LiDAR scene completion with high-quality results by reducing sampling steps and incorporating a Structural Loss to maintain geometric structure.

diffusion models3D LiDAR scene completiontraining stabilitysampling speedautonomous vehiclesdistillation methodStructural Lossscene-wise termpoint-wise termSemanticKITTI

Abstract

Diffusion models have been applied to 3D LiDAR scene completion due to their strong training stability and high completion quality. However, the slow sampling speed limits the practical application of diffusion-based scene completion models since autonomous vehicles require an efficient perception of surrounding environments. This paper proposes a novel distillation method tailored for 3D LiDAR scene completion models, dubbed ScoreLiDAR, which achieves efficient yet high-quality scene completion. ScoreLiDAR enables the distilled model to sample in significantly fewer steps after distillation. To improve completion quality, we also introduce a novel Structural Loss, which encourages the distilled model to capture the geometric structure of the 3D LiDAR scene. The loss contains a scene-wise term constraining the holistic structure and a point-wise term constraining the key landmark points and their relative configuration. Extensive experiments demonstrate that ScoreLiDAR significantly accelerates the completion time from 30.55 to 5.37 seconds per frame (>5times) on SemanticKITTI and achieves superior performance compared to state-of-the-art 3D LiDAR scene completion models. Our code is publicly available at https://github.com/happyw1nd/ScoreLiDAR.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号