Alina Roitberg
収録論文 20本 ・ フィジカルAI/ロボット学習
VLA人物-物体インタラクション検出行動認識
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- 産業用物体検出における人間可読なXAIのための視覚言語モデル適応VLA2026/9/17
産業製造ラインの物体検出において、非専門家向けに直感的な説明を生成する視覚言語モデルを微調整し、XAIインターフェースを構築した。
- RoHOI: 人物-物体インタラクション検出のためのロバスト性ベンチマーク人物-物体インタラクション検出2025/7/1
人物-物体インタラクション検出モデルのロバスト性を評価する初のベンチマークRoHOIを提案し、20種類の汚染条件下での性能低下を分析。ロバスト性向上のためのSAMPL戦略も提案。
- 指示付き原子行動認識行動認識2024/7/1
テキスト記述に基づいて特定人物の原子行動を動画から認識する新タスクRAVARを提案し、データセットRefAVAと専用手法RefAtomNetを構築した。
- ノイズラベルに頑健な骨格ベース人間行動認識フレームワークNoiseEraSAR行動認識2024/3/1
骨格データによる行動認識で問題となるラベルノイズに対処するため、サンプル選択・共教示・クロスモーダルMoEを統合したNoiseEraSARを提案し、ベンチマークで最高性能を達成した。
- Navigating Open Set Scenarios for Skeleton-based Action Recognition2023/12/1
- Quantized Distillation: Optimizing Driver Activity Recognition Models for Resource-Constrained Environments2023/11/1
- Exploring Self-supervised Skeleton-based Action Recognition in Occluded Environments2023/9/1
- Elevating Skeleton-Based Action Recognition with Efficient Multi-Modality Self-Supervision2023/9/1
- Exploring Few-Shot Adaptation for Activity Recognition on Diverse Domains2023/5/1
- Towards Activated Muscle Group Estimation in the Wild2023/3/1
- FishDreamer: Towards Fisheye Semantic Completion via Unified Image Outpainting and Segmentation2023/3/1
- Multi-modal Depression Estimation based on Sub-attentional Fusion2022/7/1
- Towards Robust Semantic Segmentation of Accident Scenes via Multi-Source Mixed Sampling and Meta-Learning2022/3/1
- TransDARC: Transformer-based Driver Activity Recognition with Latent Space Feature Calibration2022/3/1
- Delving Deep into One-Shot Skeleton-based Action Recognition with Diverse Occlusions2022/2/1
- TransKD: Transformer Knowledge Distillation for Efficient Semantic Segmentation2022/2/1
- Transfer beyond the Field of View: Dense Panoramic Semantic Segmentation via Unsupervised Domain Adaptation2021/10/1
- DensePASS: Dense Panoramic Semantic Segmentation via Unsupervised Domain Adaptation with Attention-Augmented Context Exchange2021/8/1
- Let's Play for Action: Recognizing Activities of Daily Living by Learning from Life Simulation Video Games2021/7/1
- MASS: Multi-Attentional Semantic Segmentation of LiDAR Data for Dense Top-View Understanding2021/7/1