日本フィジカルAI新聞

世界のフィジカルAIを、日本語で。

週刊ニュースレター購読
arXiv:2507.06687

StixelNExT++: Lightweight Monocular Scene Segmentation and Representation for Collective Perception

StixelNExT++: Lightweight Monocular Scene Segmentation and Representation for Collective Perception

シェア:XThreadsFacebookLINEはてブBluesky

著者: Marcel Vosshans, Omar Ait-Aider, Youcef Mezouar, Markus Enzweiler

分類: cs.CV, cs.RO

原文アブストラクト

This paper presents StixelNExT++, a novel approach to scene representation for monocular perception systems. Building on the established Stixel representation, our method infers 3D Stixels and enhances object segmentation by clustering smaller 3D Stixel units. The approach achieves high compression of scene information while remaining adaptable to point cloud and bird's-eye-view representations. Our lightweight neural network, trained on automatically generated LiDAR-based ground truth, achieves real-time performance with computation times as low as 10 ms per frame. Experimental results on the Waymo dataset demonstrate competitive performance within a 30-meter range, highlighting the potential of StixelNExT++ for collective perception in autonomous systems.