K. Madhava Krishna
収録論文 67本 ・ フィジカルAI/ロボット学習
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- 不整地における遮蔽を考慮した準静的・安定性指向の軌道計画走行計画2026/9/30
地形の遮蔽による不確かさを考慮し、不整地を走行する四輪車の安定性を重視した軌道を生成するモデルベース手法を提案した。
- GlassFormer: レーダーと深度の融合によるリアルタイムガラスセグメンテーションの学習マルチモーダル認識2026/9/29
ミリ波レーダーとRGB-Dを融合し、視覚や深度が苦手な透明ガラス面をリアルタイムにセグメンテーションする軽量トランスフォーマーネットワークを提案。
- SILICA: 拡散モデルの事前知識を再利用したガラス領域分割と深度推定の統合深度推定/セグメンテーション2026/7/1
テキストから画像を生成する拡散モデルの事前知識を活用し、ガラスの領域分割と深度推定を同時に行う統一パイプラインを提案。実世界のガラス深度アノテーションを不要とし、ゼロショットで未知環境に一般化できることを示した。
- 予測・適応・行動:タスク計画のためのハイブリッドフレームワークタスク計画2026/2/1
LLMの予測能力と確率的逐次意思決定を組み合わせ、人間の能力不足や物体欠如による失敗を予測・防止・回復するロボットタスク計画手法を提案。
- Crowd-FM: 群集ナビゲーションのための条件付きフローマッチング生成軌道の学習に基づく最適選択群集ナビゲーション2026/2/1
条件付きフローマッチングで衝突回避軌道を生成し、人間らしさを評価するスコア関数で最適な軌道を選択する群集ナビゲーション手法を提案。
- MonoMPC: 学習した衝突モデルとリスク対応型モデル予測制御による単眼視ナビゲーションナビゲーション2025/8/1
単眼RGBカメラからの推定深度を直接衝突判定に使わず、学習した衝突モデルへの入力として活用し、リスク対応型MPCで未知環境をナビゲーションする手法を提案。
- Diffusion-FS: 拡散モデルによる自動運転のためのマルチモーダル自由空間予測自動運転/自由空間予測2025/7/1
単眼カメラ画像から自己教師あり学習で走行可能な自由空間の輪郭を生成し、拡散モデルで予測する手法を提案。
- SparseLoc: 自律ナビゲーションのためのスパースオープンセットランドマークベースのグローバル位置推定位置推定2025/3/1
視覚言語基盤モデルを活用してスパースな意味トポメトリック地図をゼロショットで生成し、モンテカルロ位置推定と後期最適化を組み合わせることで、高精度かつ効率的なグローバル位置推定を実現した。
- Swarm-Gen: 多様で実行可能な群行動の高速生成群制御2025/1/1
生成モデルと安全フィルタを組み合わせ、数十ミリ秒で多様かつ実行可能な群ロボットの軌道を生成する手法を提案。
- Imagine2Servo: 拡散モデルによる目標画像生成を活用した知的ビジュアルサーボビジュアルサーボ2024/10/1
拡散モデルで中間目標画像を生成し、目標画像が不要で初期・目標画像の重なりが小さい場合でも動作するビジュアルサーボ手法を提案。実機実験で長距離ナビゲーションやマニピュレーションに有効性を示した。
- CrowdSurfer: ベクトル量子化VAEとサンプリング最適化による密集群衆ナビゲーション群衆ナビゲーション2024/9/1
VQ-VAEで専門家軌道の事前分布を学習し、実行時にサンプリング最適化で洗練させることで、動的障害物の予測なしに密集群衆内の移動成功率を40%向上させた。
- 視覚言語ナビゲーションのためのオープンセット3D意味インスタンスマップVLA2024/4/1
基盤モデルを活用し、インスタンスレベルの埋め込みを持つ3D点群マップを構築することで、言語指示に基づくナビゲーションの成功率を向上させ、未知の物体も認識可能にした。
- 微分可能な車輪-地形相互作用モデルを用いた不整地における二段階軌道最適化軌道計画2024/4/1
地形の標高データのみから車輪と地形の相互作用を非線形最小二乗問題としてモデル化し、微分可能な姿勢予測を用いた二段階最適化で不整地の軌道計画を行う手法を提案した。
- LeGo-Drive: 言語強化型ゴール指向閉ループ自律運転自動運転2024/3/1
言語指示からゴール位置を推定し、ゴールと軌道を反復的に改善する閉ループ自律運転フレームワークを提案。シミュレーションで成功率81%を達成。
- ATPPNet: Attention based Temporal Point cloud Prediction Network2024/1/1
- Automated Detection and Counting of Windows using UAV Imagery based Remote Sensing2023/11/1
- Talk2BEV: Language-enhanced Bird's-eye View Maps for Autonomous Driving2023/10/1
- Hilbert Space Embedding-based Trajectory Optimization for Multi-Modal Uncertain Obstacle Trajectory Prediction2023/10/1
- NeuroSMPC: A Neural Network guided Sampling Based MPC for On-Road Autonomous Driving2023/10/1
- UAP-BEV: Uncertainty Aware Planning using Bird's Eye View generated from Surround Monocular Images2023/6/1
- Instance-Level Semantic Maps for Vision Language Navigation2023/5/1
- FinderNet: A Data Augmentation Free Canonicalization aided Loop Detection and Closure technique for Point clouds in 6-DOF separation2023/4/1
- MVRackLay: Monocular Multi-View Layout Estimation for Warehouse Racks and Shelves2022/11/1
- UAV-based Visual Remote Sensing for Automated Building Inspection2022/9/1
- Real-Time Heuristic Framework for Safe Landing of UAVs in Dynamic Scenarios2022/9/1
- Leveraging Distributional Bias for Reactive Collision Avoidance under Uncertainty: A Kernel Embedding Approach2022/8/1
- Flow Synthesis Based Visual Servoing Frameworks for Monocular Obstacle Avoidance Amidst High-Rises2022/7/1
- Drift Reduced Navigation with Deep Explainable Features2022/3/1
- ReF -- Rotation Equivariant Features for Local Feature Matching2022/3/1
- Non Holonomic Collision Avoidance of Dynamic Obstacles under Non-Parametric Uncertainty: A Hilbert Space Approach2021/12/1
- Learning Actions for Drift-Free Navigation in Highly Dynamic Scenes2021/10/1
- AutoLay: Benchmarking amodal layout estimation for autonomous driving2021/8/1
- Monocular Multi-Layer Layout Estimation for Warehouse Racks2021/3/1
- RP-VIO: Robust Plane-based Visual-Inertial Odometry for Dynamic Environments2021/3/1
- RoRD: Rotation-Robust Descriptors and Orthographic Views for Local Feature Matching2021/3/1
- Fast Adaptation of Manipulator Trajectories to Task Perturbation By Differentiating through the Optimal Solution2020/11/1
- DRACO: Weakly Supervised Dense Reconstruction And Canonicalization of Objects2020/11/1
- BirdSLAM: Monocular Multibody SLAM in Bird's-Eye View2020/11/1
- Student Mixture Model Based Visual Servoing2020/6/1
- SROM: Simple Real-time Odometry and Mapping using LiDAR data for Autonomous Vehicles2020/5/1
- Reconstruct, Rasterize and Backprop: Dense shape and pose estimation from a single image2020/4/1
- DFVS: Deep Flow Guided Scene Agnostic Image Based Visual Servoing2020/3/1
- Multi-object Monocular SLAM for Dynamic Environments2020/2/1
- MonoLayout: Amodal scene layout from a single image2020/2/1
- Topological Mapping for Manhattan-like Repetitive Environments2020/2/1
- Reactive Navigation under Non-Parametric Uncertainty through Hilbert Space Embedding of Probabilistic Velocity Obstacles2020/1/1
- A Hierarchical Network for Diverse Trajectory Proposals2019/6/1
- IVO: Inverse Velocity Obstacles for Real Time Navigation2019/5/1
- Learning to Prevent Monocular SLAM Failure using Reinforcement Learning2018/12/1
- Parameter Sharing Reinforcement Learning Architecture for Multi Agent Driving Behaviors2018/11/1
- Solving Chance Constrained Optimization under Non-Parametric Uncertainty Through Hilbert Space Embedding2018/11/1
- Gradient Aware - Shrinking Domain based Control Design for Reactive Planning Frameworks used in Autonomous Vehicles2018/4/1
- Geometric Consistency for Self-Supervised End-to-End Visual Odometry2018/4/1
- MergeNet: A Deep Net Architecture for Small Obstacle Discovery2018/3/1
- Model Predictive Control for Autonomous Driving considering Actuator Dynamics2018/3/1
- CalibNet: Geometrically Supervised Extrinsic Calibration using 3D Spatial Transformer Networks2018/3/1
- The Earth ain't Flat: Monocular Reconstruction of Vehicles on Steep and Graded Roads from a Moving Camera2018/3/1
- Constructing Category-Specific Models for Monocular Object-SLAM2018/2/1
- Beyond Pixels: Leveraging Geometry and Shape Cues for Online Multi-Object Tracking2018/2/1
- Model Predictive Control for Autonomous Driving Based on Time Scaled Collision Cone2017/12/1
- CObRaSO: Compliant Omni-Direction Bendable Hybrid Rigid and Soft OmniCrawler Module2017/9/1
- Exploring Convolutional Networks for End-to-End Visual Servoing2017/6/1
- Design and optimal springs stiffness estimation of a Modular OmniCrawler in-pipe climbing Robot2017/6/1
- COCrIP: Compliant OmniCrawler In-pipeline Robot2017/4/1
- Reconstructing Vechicles from a Single Image: Shape Priors for Road Scene Understanding2016/9/1
- Chance constraint based multi agent navigation under uncertainty2016/8/1
- Learning to Prevent Monocular SLAM Failure using Reinforcement Learning2016/7/1