Subramanian Ramamoorthy
University of Edinburgh
収録論文 66本 ・ フィジカルAI/ロボット学習
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- フローマッチングVLAにおけるタスク依存計算配分のための分離型アーリーイグジットVLA2026/9/24
フローマッチングVLAに軽量なExit Transformerを追加し、バックボーン・行動エキスパート・デノイジングの深さをタスクごとに調整可能にした。
- 意味的制約下での未知部品を用いた新規構造の組み立て学習組み立て学習2026/8/13
訓練時には存在しなかった意味的制約(どの部品タイプや特徴が有効な構造を作るか)を、対話とデモンストレーションから学習し、新しい構造を組み立てるニューロシンボリックアーキテクチャを提案した。
- RENEW: 選好から世界モデルの学習とモデル悪用の修復を目指してモデルベースRL2026/7/15
オフライン強化学習の世界モデルがデータの薄い領域で幻覚を起こす問題に対し、人間の選好を利用してモデルのダイナミクスを直接修正する手法を提案した。不確実性に基づく効率的な微調整により、サンプル効率を改善し、破滅的忘却を抑える。
- BayesContact: 視覚触覚プロポーザルとシミュレーションベース推論による不確実な姿勢推定マニピュレーション2026/7/1
ペグ穴挿入タスクにおいて、視覚と触覚の情報をシミュレーションベース推論で融合し、物体姿勢の確率分布を推定する手法を提案。実ロボット実験で挿入成功率を30%向上させた。
- 不完全な世界モデルは悪用可能である強化学習2026/5/15
強化学習における世界モデルの悪用可能性を定義し、報酬ハッキングとの理論的関係を明らかにした論文。
- 布を広げる動作を学習:変形可能物体操作へのワールドモデルのスケーリング変形可能物体操作2026/2/1
DreamerV2を改良し、表面法線入力を用いたワールドモデルで空中での布広げ操作を学習。シミュレーションと実機でのゼロショット展開で汎化性能を示した。
- 確率的動的システムにおける尤度不要推論のための潜在的に誤指定されたドメインサポートのヒューリスティック適応sim2real2025/10/1
ロボット工学における尤度不要推論(LFI)において、サポートの誤指定問題を解決するため、EDGE、MODE、CENTREという3つのヒューリスティックLFI変種を提案し、動的変形可能線形物体(DLO)操作タスクで有効性を示した。
- ROOM: フォトリアリスティックな医療データセット生成のための物理ベース連続体ロボットシミュレータsim2real2025/9/1
患者CTスキャンを用いて気管支鏡検査のフォトリアルなマルチモーダルセンサデータを生成する物理ベース連続体ロボットシミュレータROOMを提案し、姿勢推定や深度推定タスクでその有効性を検証した。
- Assistax: 支援ロボティクス向けマルチエージェント・ハードウェア加速強化学習ベンチマークマルチエージェント強化学習2025/7/1
JAXとMuJoCo-MJX上でGPU加速された支援ロボティクスタスク群を提供し、人間役のヒューマノイドエージェントとロボットをマルチエージェント強化学習で共訓練できるベンチマークを構築した。
- ContactFusion: 視覚と接触センシングを統合した確率的ポアソン表面マップマニピュレーション2025/3/1
力覚センサからエンドエフェクタ上の接触位置を推定し、視覚情報と統合して確率的ポアソン表面マップを構築することで、ペグ挿入タスクの穴位置推定精度を向上させた研究。
- 視覚駆動型変形可能線状物体操作のための物体中心エージェント適応におけるReal2Sim2Realの分布的アプローチsim2real2025/2/1
変形可能線状物体(DLO)の操作において、尤度フリー推論で物理パラメータの事後分布を推定し、ドメインランダム化を用いたシミュレーション訓練で物体固有の視覚運動ポリシーを学習、実世界でゼロショット展開する統合フレームワークを提案。
- 対話によるコード生成:自動運転車テスト用シナリオ生成のための対話システム設計事例VLA2024/10/1
自動運転車のシミュレーションテスト用に、自然言語でシナリオを生成する対話システムを設計。対話が成功に重要で、成功率が4.5倍向上することを示した。
- 非線形力学系を内包するリザバー層による模倣学習模倣学習2024/9/1
リザバーコンピューティングに着想を得て、固定の非線形力学系を組み込んだRNN層を提案し、模倣学習における誤差蓄積を抑制して手書き動作の再現精度と頑健性を向上させた。
- SECURE: 未知の概念に対応する意味論的対話型ロボット学習対話型学習2024/9/1
ロボットが知らない概念を含む指示(例:「グラニースミスをカゴに入れて」)に対し、対話と修正フィードバックを通じて概念を学習し、タスクを遂行する手法を提案。
- 希少な交通違反に対する再利用可能な時間モニタの適応的分割sim2real2024/5/1
自動運転車の安全性検証において、希少な違反確率を効率的に推定するため、適応的多段階分割とSTLロバストネス指標を組み合わせ、部分軌跡の計算を再利用する手法を提案した。
- DROID: 大規模実環境ロボットマニピュレーションデータセットマニピュレーション2024/3/1
北米・アジア・欧州の50名が12ヶ月かけて564シーン・84タスクで収集した76k軌道・350時間のロボット操作データセットを構築し、学習ポリシーの性能と汎化性能が向上することを示した。
- クリックで掴む:視覚拡散記述子によるゼロショット精密マニピュレーションマニピュレーション2024/3/1
テキスト画像拡散モデルを活用し、ユーザーが指定した画像上のクリック位置に対応する対象物の部位を別シーンでゼロショットに把持する手法を提案。
- Open X-Embodiment: Robotic Learning Datasets and RT-X Models2023/10/1
- Generating robotic elliptical excisions with human-like tool-tissue interactions2023/9/1
- Comparison of Pedestrian Prediction Models from Trajectory and Appearance Data for Autonomous Driving2023/5/1
- DiPA: Probabilistic Multi-Modal Interactive Prediction for Autonomous Driving2022/10/1
- Learning robotic cutting from demonstration: Non-holonomic DMPs using the Udwadia-Kalaba method2022/9/1
- Testing Rare Downstream Safety Violations via Upstream Adaptive Sampling of Perception Error Models2022/9/1
- Perspectives on the System-level Design of a Safe Autonomous Driving Stack2022/8/1
- A Novel Design and Evaluation of a Dactylus-Equipped Quadruped Robot for Mobile Manipulation2022/7/1
- On Specifying for Trustworthiness2022/6/1
- Achieving Dexterous Bidirectional Interaction in Uncertain Conditions for Medical Robotics2022/6/1
- Risk-Driven Design of Perception Systems2022/5/1
- Flash: Fast and Light Motion Prediction for Autonomous Driving with Bayesian Inverse Planning and Learned Motion Profiles2022/3/1
- Learning physics-informed simulation models for soft robotic manipulation: A case study with dielectric elastomer actuators2022/2/1
- Automated Testing with Temporal Logic Specifications for Robotic Controllers using Adaptive Experiment Design2021/9/1
- Attainment Regions in Feature-Parameter Space for High-Level Debugging in Autonomous Robots2021/8/1
- Interpretable Goal Recognition in the Presence of Occluded Factors for Autonomous Vehicles2021/8/1
- Learning Time-Invariant Reward Functions through Model-Based Inverse Reinforcement Learning2021/7/1
- Formation Control for UAVs Using a Flux Guided Approach2021/3/1
- PILOT: Efficient Planning by Imitation Learning and Optimisation for Safe Autonomous Driving2020/11/1
- ProbRobScene: A Probabilistic Specification Language for 3D Robotic Manipulation Environments2020/11/1
- Affordance-Aware Handovers with Human Arm Mobility Constraints2020/10/1
- Counterfactual Explanation and Causal Inference in Service of Robustness in Robot Control2020/9/1
- Action sequencing using visual permutations2020/8/1
- Residual Learning from Demonstration: Adapting DMPs for Contact-rich Manipulation2020/8/1
- Self-Assessment of Grasp Affordance Transfer2020/7/1
- Semi-supervised Learning From Demonstration Through Program Synthesis: An Inspection Robot Case Study2020/7/1
- From Demonstrations to Task-Space Specifications: Using Causal Analysis to Extract Rule Parameterization from Demonstrations2020/6/1
- Learning from Demonstration with Weakly Supervised Disentanglement2020/6/1
- Affordances in Robotic Tasks -- A Survey2020/4/1
- Interpretable Goal-based Prediction and Planning for Autonomous Driving2020/2/1
- Elaborating on Learned Demonstrations with Temporal Logic Specifications2020/2/1
- A Two-Stage Optimization-based Motion Planner for Safe Urban Driving2020/2/1
- Learning rewards for robotic ultrasound scanning using probabilistic temporal ranking2020/2/1
- Learning Structured Representations of Spatial and Interactive Dynamics for Trajectory Prediction in Crowded Scenes2019/11/1
- Surfing on an uncertain edge: Precision cutting of soft tissue using torque-based medium classification2019/9/1
- Hybrid system identification using switching density networks2019/7/1
- Exploiting Causality for Selective Belief Filtering in Dynamic Bayesian Networks (Extended Abstract)2019/7/1
- Composing Diverse Policies for Temporally Extended Tasks2019/7/1
- Vid2Param: Modelling of Dynamics Parameters from Video2019/7/1
- Disentangled Relational Representations for Explaining and Learning from Demonstration2019/7/1
- Learning Grasp Affordance Reasoning through Semantic Relations2019/6/1
- DynoPlan: Combining Motion Planning and Deep Neural Network based Controllers for Safe HRL2019/6/1
- Reasoning on Grasp-Action Affordances2019/5/1
- To Stir or Not to Stir: Online Estimation of Liquid Properties for Pouring Actions2019/4/1
- From explanation to synthesis: Compositional program induction for learning from demonstration2019/2/1
- Active Localization of Gas Leaks using Fluid Simulation2019/1/1
- Interpretable Latent Spaces for Learning from Demonstration2018/7/1
- FPR -- Fast Path Risk Algorithm to Evaluate Collision Probability2018/4/1
- The ORCA Hub: Explainable Offshore Robotics through Intelligent Interfaces2018/3/1