SIREN-Bench: 行動駆動型の緊急車両インタラクション生成と評価
SIREN-Bench: Behavior-Driven Generation and Evaluation of Emergency-Vehicle Interactions
緊急車両と一般車両のインタラクションを行動レベルで制御できるSUMO-CARLA共シミュレーションプラットフォームSIRENを提案し、ベンチマークとして3D物体検出、軌道予測、視覚言語リスク理解のタスクで評価した。
詳しい要約
1. どんなもの?
2. 先行研究と比べてどこがすごい?
3. 技術・手法の肝は?
4. どうやって有効だと検証した?
5. 議論はある?
6. 次に読むべき論文は?
※ AIが要旨から生成した要約です。正確性は原文をご確認ください。
著者: Yicheng Zhu, Tianmu Zhao, Haoxin Leng, Fan Zuo, Tao Li, Zilin Bian
分類: cs.RO
原文アブストラクト
Emergency vehicles (EMVs) can reorganize surrounding traffic as civilian vehicles brake, change lanes, or form rescue corridors in response to their passage. Evaluating these safety-critical interactions requires behavior-level control over both EMV privileges and civilian responses, together with consistent sensing and ground truth. Existing datasets and simulation benchmarks do not directly provide this combination. We present \textbf{SIREN}, a behavior-driven SUMO--CARLA co-simulation platform for generating EMV--civilian interactions. SIREN couples SUMO's network-level traffic evolution and behavior logic with CARLA's continuous vehicle control and synchronized onboard sensing; depending on the active behavior, the interaction is controlled by SUMO, CARLA, or jointly. We instantiate the platform as \textbf{SIREN-Bench-v1}, comprising seven parameterized interaction templates across emergency levels L1--L3 and three behavior families, with synchronized sensor observations and simulator-native annotations. We demonstrate the benchmark through three representative tasks: 3D object detection, trajectory prediction, and vision-language risk understanding. Evaluations of nine trajectory predictors, four LiDAR-based detectors, and five vision-language models reveal behavior-dependent failure modes. Traffic-clearance interactions are hardest for detection, privileged intersection traversal is hardest for prediction, and no learned predictor outperforms the constant-velocity reference on average. Vision-language models perform substantially better on normal traffic than on near-miss and collision events. These results demonstrate the value of behavior-centered benchmarking and establish SIREN as an extensible data-generation and evaluation platform for autonomous-driving and transportation safety research.
関連論文
- AM-Bench: 空中マニピュレーションのポリシー学習のためのモジュール式シミュレーションスイートとベンチマークシミュレーション/ベンチマーク
- R2HandoverSim: ロボットから人間への物体受け渡しのためのシミュレーションフレームワークとベンチマークシミュレーション/ベンチマーク
- Labimus: 化学実験室における人型器用操作のためのシミュレーションとベンチマークシミュレーション/ベンチマーク