論文
フィジカルAI関連の最新論文を、日本語のひとことで。原題は英語です。
- VLA
運転用視覚言語行動モデルにおける計画トークンの深さ方向の探索と刈り込み
運転用VLAモデルで、計画トークンが各デコーダ層でどの程度の情報を持つかを調べ、初期層で意味情報が線形分離可能であることを示し、後半層を刈り込んでも性能が保てることを発見した。
原題: Depth-Wise Probing and Pruning of the Planning Token in a Driving Vision-Language-Action Model / Harisankar Babu, Benjamin Coors, Christopher Lang 他
日本語で読む → - ロボティクス
Representation Handoffs for OpenArm-Based Laboratory Mobile Manipulation
原題: Representation Handoffs for OpenArm-Based Laboratory Mobile Manipulation / Yang Shen, Chonghao Cheng, Ziyi Zhao 他
日本語で読む → - 画像認識
Vernata: Self-Supervised Learning of LiDAR Point Representations
原題: Vernata: Self-Supervised Learning of LiDAR Point Representations / Oliver Lemke, Alexander Liniger, Abel Gawel 他
日本語で読む → - 模倣学習/可変インピーダンス制御
VIDP: 多様なデモから学ぶコンプライアントロボット操作のための可変インピーダンス拡散ポリシー
力センサを使わずに、多様なデモから可変インピーダンス制御を学習するフレームワークを提案し、実機実験で固定インピーダンス法より高い成功率と低い接触力を達成した。
原題: VIDP: Variable Impedance Diffusion Policy for Compliant Robot Manipulation from Diverse Demonstrations / Hisham Khalil, Neil Fernandes, Thomas M. Kwok 他
日本語で読む → - sim2real
R2S-EGO: スパースキャプチャ実世界からシミュレーションへのデュアルプロキシ精緻化
ロボットの軌跡に沿った視点を効率的に補完するため、シミュレータ由来のロボットプロキシとキャプチャ由来の幾何プロキシを組み合わせ、実画像をアンカーに疑似観測を生成してシーンを精緻化する手法を提案。
原題: R2S-EGO: Dual-Proxy Refinement for Sparse-Capture Real-to-Sim / Shuai Fang, Xin Deng, Yuchen Kang 他
日本語で読む → - ロボティクス
When Coordination Becomes a Threat: Communication Attacks in LLM-Controlled Multi-Robot Systems
原題: When Coordination Becomes a Threat: Communication Attacks in LLM-Controlled Multi-Robot Systems / Zhen Huang, Zhihuang Liu, Weijia Shi 他
日本語で読む → - 動作計画
任意形状ペイロードを運搬する移動マニピュレータのためのリアルタイム全身動作計画:運動学的に結合したSVSDFによるアプローチ
移動マニピュレータが任意形状の大型ペイロードを運搬する際のリアルタイム全身動作計画フレームワークを提案。チェーン分解カーネル衝突チェックと運動学的に結合したSVSDF最適化により、複雑な環境での効率的な経路生成を実現した。
原題: Real-time Whole-Body Motion Planning for Mobile Manipulators Carrying Arbitrarily Shaped Payloads via Kinematically-Coupled SVSDF / Yisheng Li, Longji Yin, Tingrui Zhang 他
日本語で読む → - ロボティクス
CrossTracer: Cross-Embodiment Navigation via VLA Model Reasoning and Trace Residuals Adapting
原題: CrossTracer: Cross-Embodiment Navigation via VLA Model Reasoning and Trace Residuals Adapting / Yao Wang, Siyuan Wang, Zhirui Sun 他
日本語で読む → - ロボティクス
Benchmarking and Reasoning Distillation of Large Language Models for Feedback Controller Design in Complex Dynamical Systems
原題: Benchmarking and Reasoning Distillation of Large Language Models for Feedback Controller Design in Complex Dynamical Systems / Zhongchao Zhou, Yixuan Xie, Wenwei Yu 他
日本語で読む → - ロボティクス
Spatiotemporal Agility: Time-Constrained Reinforcement Learning for Vision-Guided Dynamic Quadrupedal Interception
原題: Spatiotemporal Agility: Time-Constrained Reinforcement Learning for Vision-Guided Dynamic Quadrupedal Interception / Yidong Zhu, Zibo Dai, Tongning Zhang 他
日本語で読む → - シーングラフ
Prior-SG: タスクと事前知識に基づく任意構造環境におけるシーングラフの領域分割
任意構造の環境で、視覚・幾何・物体情報とLLMが生成する事前グラフを確率的に統合し、シーングラフの領域分割を高精度に行うフレームワークを提案した。
原題: Prior-SG: Task and Prior Driven Region Segmentation for Scene Graphs in Arbitrarily-Structured Environments / Giorgio Tonetti, Laurent Kneip, Abel Gawel 他
日本語で読む → - ロボティクス
C2Dex: Contact-Consistent Reconstruction and Retargeting for Dexterous Manipulation from Monocular Video
原題: C2Dex: Contact-Consistent Reconstruction and Retargeting for Dexterous Manipulation from Monocular Video / Jie Ren, Zhehao Jiang, Yinhong Yang 他
日本語で読む → - VLA/ナビゲーション
WNM-3D: 3Dシーン条件付けによるクローズドループVLNのための世界ナビゲーションモデル
連続的な視覚言語ナビゲーション(VLN)のための生成的世界行動モデルを提案し、3Dシーン表現を条件として将来の視覚と行動を同時生成することで、クローズドループのナビゲーション性能を向上させた。
原題: WNM-3D: A World Navigation Model with 3D Scene Conditioning for Closed-Loop VLN / Yuehao Huang, Yunzi Wu, Xiaotao Zhang 他
日本語で読む → - VLA
AtlasVLA: 視覚言語行動モデルのための持続的な世界・自我状態モデリング
単眼カメラのみで長期的なタスクを遂行できるよう、視覚言語行動モデルに持続的な世界状態メモリと自我作業メモリを導入し、空間的推論を強化したフレームワークを提案した。
原題: AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models / Guiyu Zhao, Longteng Guo, Yanghong Mei 他
日本語で読む → - ソーシャルロボット/評価フレームワーク
ロボット用基盤モデルはどう選ぶべきか?ソーシャルロボットのためのコミュニティ評価フレームワークを支持する
ソーシャルロボット向け基盤モデルの選択を支援するため、5つの評価次元と3段階の評価パラダイムを提案し、コミュニティでの共同構築を呼びかける論文。
原題: How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots / Eric Nichols, Alva Markelius, Hatice Gunes
日本語で読む → - ロボティクス
M2-SMap: Memory-Efficient Semantic Mapping with Hierarchical Multi-Model Representation
原題: M2-SMap: Memory-Efficient Semantic Mapping with Hierarchical Multi-Model Representation / QiYing Deng, ZhongLai Wang, Yuan Gao 他
日本語で読む → - ロボティクス
Decoupling Intention from Trajectory: A Representational Deduction Framework for World Action Models
原題: Decoupling Intention from Trajectory: A Representational Deduction Framework for World Action Models / Xiangkai Ma, Yue Ma, Junjie Wang 他
日本語で読む → - ロボティクス
Cross-View Action Consistency for Camera-Robust Vision-Language-Action Policies
原題: Cross-View Action Consistency for Camera-Robust Vision-Language-Action Policies / Bingqi Huang, Bingchuan Wei, Xuan Wang 他
日本語で読む → - ロボティクス
TEMPO: Semantic-Action Decoupled RL Post-Training for Vision-Language-Action Models
原題: TEMPO: Semantic-Action Decoupled RL Post-Training for Vision-Language-Action Models / Ziheng Liu, Quantao Yang
日本語で読む → - ロボティクス
Unordered Landmark Visual Navigation
原題: Unordered Landmark Visual Navigation / Hao Ren, Junzhe Zhu, Yihan Li 他
日本語で読む → - 触覚
6次元動的触覚センシングに基づく過渡的外部接触の検出と測距
把持物体と環境との過渡的な接触を、小型6D慣性センサを用いた動的触覚センシングで高速・高精度に検出・位置推定する手法を提案した。
原題: Detection and Ranging of Transient Extrinsic Contacts Based on 6D Dynamic Tactile Sensing / Haowen Zheng, Yinghao Wu, Fuyuan Liu 他
日本語で読む → - ロボティクス
Exact Thrust-Reversal Limits of Bidirectional Propellers under Bounded Motor Inputs
原題: Exact Thrust-Reversal Limits of Bidirectional Propellers under Bounded Motor Inputs / Ahmed Ali, Chiara Gabellieri, Antonio Franchi
日本語で読む → - ロボティクス
Identifying the Key Biomechanical Features of Movement Adaptation during Exoskeleton-Assisted Locomotion
原題: Identifying the Key Biomechanical Features of Movement Adaptation during Exoskeleton-Assisted Locomotion / Peter Seungjune Lee, Katja Mombaur
日本語で読む → - ロボティクス
A Haptic Robot Finger Designed for Guqin Instrument Playing
原題: A Haptic Robot Finger Designed for Guqin Instrument Playing / Tianwei Zhang, Hanming Yan, Yang Yang. Ziya Wang
日本語で読む → - 模倣学習/介入
AutoIntervene: アクションチャンキング模倣学習ポリシーのための校正された介入
アクションチャンキング模倣学習ポリシーがデモ分布から逸脱した際に、オペレータへの制御移行を選択的に行うオンラインフレームワークを提案。視覚・行動サポートメモリと校正された閾値で介入を制御し、実世界の双腕操作タスクで成功率を向上させた。
原題: AutoIntervene: Calibrated Intervention for Action-Chunking Imitation Learning Policies / Jinhe Tang, Weiming Zhi
日本語で読む → - ワールドモデル
前方予測だけで十分か?JEPAワールドモデルのための物理状態接地
JEPAベースのワールドモデルに、ロボットの自己受容状態と関節角変化を接地する2つの目的を追加し、潜在表現の識別性と下流タスク性能を向上させる手法を提案した。
原題: Is Forward Prediction Enough? Physical State Grounding for JEPA World Models / Haodong Yan, Jiaguan Zhu, Mingyuan Jia 他
日本語で読む → - ソフトロボティクス
SoRoMoX: 高速・微分可能・並列化可能なソフトロボットモデル
ソフトロボットの制御向けに、JAXベースで微分可能かつGPU並列実行可能なCosseratロッド理論に基づくモデリングフレームワークを開発し、従来比で最大234.6倍のスループット向上を実現した。
原題: SoRoMoX: Fast, Differentiable, and Parallelizable Soft Robot Models / Maximilian Stölzle, Solange Gribonval, Daniel Feliu-Talegon 他
日本語で読む → - ロボティクス
Hoverflie: An empirical investigation of rotor shrouds to transform micro air vehicles into multi-modal hovercraft
原題: Hoverflie: An empirical investigation of rotor shrouds to transform micro air vehicles into multi-modal hovercraft / Mrinmoy Modak, Daniel S. Drew
日本語で読む → - 群制御
SyncSBC: 同期自律制御のための分散型群行動予測
本論文では、各エージェントが局所的な知覚のみから群全体の行動を分類し、分散合意により群の意思決定を同期させる手法SyncSBCを提案し、実ロボットでの異常検知と行動変化の自律調整を実証した。
原題: SyncSBC: Decentralized Swarm Behavior Prediction for Synchronized Autonomous Control / Varun Raveendra, Connor Mattson, Daniel S. Brown
日本語で読む → - cs.ET
Ising Acceleration for Multi-Robot Multi-Target Planning
原題: Ising Acceleration for Multi-Robot Multi-Target Planning / Ahmet Efe, Recep B. Uludag, Chris H. Kim 他
日本語で読む → - ロボティクス
Learning Fault-Tolerant Locomotion with Adaptive Gait Timing
原題: Learning Fault-Tolerant Locomotion with Adaptive Gait Timing / Giovanbattista Gravina, Luca Rossini, Carlo Rizzardo 他
日本語で読む → - ロボティクス
Acoustic-driven millimetric helical robot: ultrasonic synergistic manipulation in confined fluidic environment
原題: Acoustic-driven millimetric helical robot: ultrasonic synergistic manipulation in confined fluidic environment / Hanlin Wang, Xin Wang, Xinwei Wei 他
日本語で読む → - ロボティクス
A Disturbance in the Force: Force Actuation on the RAVEN II Surgical Robot with Parallel Motor-Cable Units
原題: A Disturbance in the Force: Force Actuation on the RAVEN II Surgical Robot with Parallel Motor-Cable Units / Haonan Peng, Dun-Tin Chiang, Jordan Hendricks 他
日本語で読む → - ロボティクス
LyEvO: Lyapunov-Guided Evolutionary Optimization for Safe and Robust Sim-to-Real Policy Learning
原題: LyEvO: Lyapunov-Guided Evolutionary Optimization for Safe and Robust Sim-to-Real Policy Learning / Riccardo Curcio, Hongpeng Cao, Marco Caccamo
日本語で読む → - 群制御
探索支援型のエージェント・環境協調強化学習による回転動作を考慮した頑健な生涯マルチエージェント経路探索
実世界の倉庫システムを模した回転制約と安全制約を持つ生涯マルチエージェント経路探索問題に対し、ニューラル方策と探索ベースのプランナーを組み合わせ、エージェントと環境の方策を同時に学習する手法を提案した。
原題: Search-Aided Joint Agent-Environment Reinforcement Learning for Robust Lifelong Multi-Agent Path Finding with Rotations / He Jiang, Jingtian Yan, Yulun Zhang 他
日本語で読む → - 操作
SkillMemo: 専門家ガイドによるスキル記憶フレームワークを用いた構成可能な身体操作
長期的なデモンストレーションを潜在的な原子スキルに分解し、動的エピソード記憶バンクに統合することで、構成可能な操作タスクの汎化を向上させるフレームワークを提案した。
原題: SkillMemo: Expert-guided Skill Memory Framework for Compositional Embodied Manipulation / Changyuan Wang, Chubin Zhang, Zhenyu Wu 他
日本語で読む → - シミュレーション評価
GAUGE: 物理的忠実性を測定するための実世界基盤ベンチマーク
シミュレーションエンジンと生成ビデオワールドモデルの物理的忠実性を、実世界の軌跡に基づいて診断するベンチマークを提案した。
原題: GAUGE: A Measurement-Grounded Benchmark for Physical Fidelity in Simulation Engines and Video World Models / Shuai Wang, Yaxin Feng, Xuekun Jiang 他
日本語で読む → - 強化学習
観測に基づく自己予測強化学習による視覚連続制御
視覚ベースの連続制御タスクにおいて、潜在空間での自己予測と観測空間での予測を組み合わせた新しい表現学習手法を提案し、サンプル効率を向上させた。
原題: Observation-Grounded Self-Predictive Reinforcement Learning for Visual Continuous Control / Xinwei Liu, Junyuan Liang, Jianting Zhang 他
日本語で読む → - ナビゲーション
PathCover: 点群に対するランダム反復空間分割による高速凸分解
点群データから障害物のない凸領域を高速に生成する新しいアルゴリズムを提案し、ロボットの軌道計画を高速化する。
原題: PathCover: A Fast Convex Decomposition along a Path via Randomized Iterative Space Partitioning (RISP) on Point Clouds / Kunal S. Narkhede, Abhijeet M. Kulkarni, Guoquan Huang 他
日本語で読む → - VLA
ARGUS: 大規模3Dビジョンモデルによる視点変化下でのロボットシーン幾何の整合
視点に依存しない観測表現を生成する前処理パイプラインを提案し、多視点データセットからの学習効率と汎化性能を向上させた。
原題: ARGUS: Aligning Robot Scene Geometry Under Shifting Views with Large 3D Vision Models / Rishik Sathua, Haonan Chen, Katherine Driggs-Campbell
日本語で読む → - 状態推定
スライディングセンサ:連続体ロボットの状態推定における設定可能な信頼度
連続体ロボットの状態推定において、センサをロボット内で縦方向に移動させることで、タスクに応じた場所の推定信頼度を調整できることを示した論文。
原題: Sliding Sensors: Configurable Confidence in State Estimation for Continuum Robots / Ella Walsh, Spencer Teetaert, Eric Diller 他
日本語で読む → - 機械学習
Is Self-Pretraining really useful to improve diagnosis in medical Time Series?
原題: Is Self-Pretraining really useful to improve diagnosis in medical Time Series? / Omar Coser, Antonio Orvieto, Paolo Soda 他
日本語で読む → - ロボティクス
KILVO: Kinematic-Inertial-LiDAR-Visual Odometry with Robust Multimodal Adaptation for Humanoid Robots
原題: KILVO: Kinematic-Inertial-LiDAR-Visual Odometry with Robust Multimodal Adaptation for Humanoid Robots / Jixin Gao, Fucheng Liu, Teng Zhang 他
日本語で読む → - ロボティクス
Design and Evaluation of a Touchscreen-Based Teleoperation Interface for Robotic Manipulators
原題: Design and Evaluation of a Touchscreen-Based Teleoperation Interface for Robotic Manipulators / Juan José García Cárdenas, Alperen Kenan, Hamidreza Raei 他
日本語で読む → - ロボティクス
JoyAI-RA 0.5: Scaling Robot Manipulation Learning via Dual Action Alignment
原題: JoyAI-RA 0.5: Scaling Robot Manipulation Learning via Dual Action Alignment / JoyAI-RA Team
日本語で読む → - ロボティクス
Robot Learning from Human Demonstrations: Handwritten Alphabet Trajectories and Human-Likeness Evaluation
原題: Robot Learning from Human Demonstrations: Handwritten Alphabet Trajectories and Human-Likeness Evaluation / Alperen Kenan, Paul Bremner, Manuel Giuliani
日本語で読む → - ロボティクス
In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use
原題: In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use / Jiarui Yang, Wen Huang, Jiale Zhang 他
日本語で読む → - ロボティクス
Fast and Accurate: An Adaptive VLA Inference Framework through Environment-aware Model Selection
原題: Fast and Accurate: An Adaptive VLA Inference Framework through Environment-aware Model Selection / Yuewei Sun, Lang Qin, Zechuan Tian 他
日本語で読む → - ロボティクス
Nonvisual Classification of Ground-Condition by Artificial Proprioception in an Amoeba-Inspired Autonomous Walking Robot
原題: Nonvisual Classification of Ground-Condition by Artificial Proprioception in an Amoeba-Inspired Autonomous Walking Robot / Hyoto Yamaguchi, Zenji Yatabe, Seiya Kasai
日本語で読む → - ロボティクス
Near-sensor Computing for Rapid Visuotactile Perception
原題: Near-sensor Computing for Rapid Visuotactile Perception / Zhengying Zhu, Ruilin Zhang, Runze Hu 他
日本語で読む → - ロボティクス
$ω$-0: A Latent Predictive World Action Model for Concurrent Humanoid Loco-Manipulation
原題: $ω$-0: A Latent Predictive World Action Model for Concurrent Humanoid Loco-Manipulation / Zhe Li, Zhenzhe Zhang, Yangyang Wei 他
日本語で読む → - ロボティクス
ErgoSurf: Ergodic Control for the Coverage of Unknown Surfaces
原題: ErgoSurf: Ergodic Control for the Coverage of Unknown Surfaces / Stefan Schneyer, Timo Bachmann, Maged Iskandar 他
日本語で読む → - ロボティクス
A Master-Slave Robot Manipulator for Needle-Based Teleoperation in MRI Chamber
原題: A Master-Slave Robot Manipulator for Needle-Based Teleoperation in MRI Chamber / Omar Curiel, Jing-Yuan Huang, Po-Chih Chen 他
日本語で読む → - ロボティクス
IcFuzz: Fuzzing Isaac Sim with Semantic Stage Guidance and Multi-level Mutation
原題: IcFuzz: Fuzzing Isaac Sim with Semantic Stage Guidance and Multi-level Mutation / Zhixiang Chen, Zhuangbin Chen, Ruoxi Jia 他
日本語で読む → - ロボティクス
Coordinated Multi-Robot Disassembly for Makespan Optimization of Large-Scale Assemblies
原題: Coordinated Multi-Robot Disassembly for Makespan Optimization of Large-Scale Assemblies / Niklas Hargus, Andreas Orthey, Marc Toussaint
日本語で読む → - ロボティクス
Hijacking Robots with a Piece of Paper: A Systematic Study of Physical Prompt Injection in VLM-Controlled Robots
原題: Hijacking Robots with a Piece of Paper: A Systematic Study of Physical Prompt Injection in VLM-Controlled Robots / S. M . Bhagya P. Samarakoon, M. A. Viraj J. Muthugala, W. K. R. Sachinthana 他
日本語で読む → - ロボティクス
Beyond Flat Policies: Hierarchical Post-Training for Embodied Agents in Robotic Manipulation
原題: Beyond Flat Policies: Hierarchical Post-Training for Embodied Agents in Robotic Manipulation / He Kong, Zengjue Chen, Qi Wang 他
日本語で読む → - ロボティクス
DyPES-VLA: Learning Shared Dynamics Priors and Embodiment-Specific Control for Cross-Embodiment Manipulation
原題: DyPES-VLA: Learning Shared Dynamics Priors and Embodiment-Specific Control for Cross-Embodiment Manipulation / Junfeng Li, Junjie He, Zhide Zhong 他
日本語で読む → - ロボティクス
TRACE: Learned Proprioceptive Odometry for Legged Robots under Unreliable Contact Conditions
原題: TRACE: Learned Proprioceptive Odometry for Legged Robots under Unreliable Contact Conditions / Taehyeon Kong, Woojin Kim, Jemin Hwangbo
日本語で読む → - ロボティクス
GeniWorld: A Generalizable Interactive World Model for Robotic Manipulation via Visual Actions
原題: GeniWorld: A Generalizable Interactive World Model for Robotic Manipulation via Visual Actions / Chenghao Gu, Hanyang Yu, Jingbo Zhang 他
日本語で読む →