Michael Milford
収録論文 108本 ・ フィジカルAI/ロボット学習
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- イベントカメラによる多視点ジオローカリゼーションVPR/イベントカメラ2026/9/18
フレームベースのデータセットをイベントストリームに変換して視点変化に頑健なイベントベースの位置認識モデルMegaEventを学習し、新規データセットSpringfield-Event-VPRを提案した。
- イベントカメラを用いたロボティクスビジョンのためのEventCVライブラリイベントカメラ/ロボティクスビジョン2026/9/18
イベントカメラをロボットに統合するためのオープンソースRustライブラリEventCVを開発し、ノイズ除去や特徴学習、ONNX推論などの機能を提供してロボット応用事例を示した。
- 既視感を打ち破る:視覚言語推論による視覚場所認識の独立監査視覚場所認識2026/7/14
視覚場所認識の結果を、視覚言語モデルを用いて独立に検証する新しい監査フレームワークを提案し、誤受理率を低減しつつ再現率を向上させた。
- DisPlace: 複数参照を用いた視覚的位置認識のための識別的場所射影視覚的位置認識2026/5/1
複数の参照走行から得た記述子を融合して、場所の識別性を高めつつ環境変化による変動を抑える新しい視覚的位置認識手法を提案した。
- 長期的な外観変化下でのマルチセッション3D再構成3D再構成/SfM2026/2/1
サンゴ礁調査のように年月を経て大きく見た目が変わる環境でも、複数回の撮影を統合して一貫した3Dモデルを再構成する手法を提案。手作り特徴と学習特徴を組み合わせ、セッション間の対応をSfM内で直接強制する。
- 分位点転送による視覚的場所認識の信頼性の高い動作点選択視覚的場所認識2026/2/1
視覚的場所認識システムの動作点(マッチング閾値)を、キャリブレーション走行と分位点正規化を用いて自動推定し、環境変化に適応させる手法を提案。
- サリエンシー誘導による左側通行自動運転のためのドメイン適応自動運転/ドメイン適応2025/11/1
米国の右側通行データで学習したPilotNetを、反転データでの事前学習と豪州高速道路データでの微調整により左側通行環境へ適応させ、ステアリング予測精度と注意領域の変化をサリエンシー解析で評価した。
- Pixi: ロボティクスとAIのための統合ソフトウェア開発・配布基盤ソフトウェア基盤2025/11/1
ロボティクス研究の再現性危機を解決するため、依存関係をロックファイルで固定し、高速SATソルバーとconda-forge/PyPI統合により環境構築を数分に短縮するパッケージ管理フレームワークPixiを提案。
- 場所認識の探求:人工システムと自然システムにおける比較場所認識2025/11/1
ロボット・動物・人間の場所認識メカニズムを比較し、共通する計算戦略や課題を整理したレビュー。
- 疑いのレンズを通して:視覚的位置認識のための堅牢で効率的な不確実性推定VPR/不確実性推定2025/10/1
視覚的位置認識(VPR)において、既存手法の類似度スコアの統計的パターンを分析するだけで、追加学習なしに予測信頼度を推定する3つの指標を提案し、9手法・6データセットで有効性を示した。
- 照明変化に頑健なイベントカメラ位置認識のためのアンサンブル手法位置認識2025/9/1
イベントカメラを用いた位置認識において、複数の再構成手法・特徴抽出器・時間解像度を組み合わせたアンサンブル手法を提案し、昼夜を問わない照明変化下での認識性能を大幅に向上させた。
- ワープ速度への準備:イベントカメラによるサブミリ秒視覚位置認識VPR/イベントカメラ2025/9/1
イベントカメラのサブミリ秒スライスからアクティブピクセル位置を二値フレームで符号化し、ビット演算で高速照合する軽量な視覚位置認識システムを提案。既存手法より大幅に高精度・低遅延を実現。
- Event-LAB: ニューロモーフィック位置推定手法の標準評価フレームワークイベントカメラ/位置推定2025/9/1
イベントベース位置推定の手法とデータセットを統一環境で実行・比較できるフレームワークEvent-LABを提案し、VPRとSLAMで有効性を示した。
- 高速フーリエ領域相互相関によるイベントカメラベースの視覚教示・再現ナビゲーション視覚ナビゲーション2025/9/1
イベントカメラのストリームマッチングを周波数領域の相互相関で高速化し、屋内外3000m以上の経路を15cm以下の誤差で自律走行できる視覚教示・再現ナビゲーションを実現した。
- VLM誘導による惑星規模の視覚的位置認識VLA2025/7/1
VLMで候補地域を絞り込み、検索ベースの視覚的位置認識と再ランキングを組み合わせることで、地球規模の画像ジオローカリゼーション精度を向上させたハイブリッド手法を提案。
- 安全なロボットナビゲーションのための視覚的位置認識における敵対的攻撃と検出VPR/敵対的攻撃/ロボットナビゲーション2025/6/1
視覚的位置認識(VPR)に対する4つの一般的な敵対的攻撃と4つのVPR特有の新規攻撃の影響を分析し、敵対的攻撃検出器(AAD)をVPRとナビゲーション判断のループに組み込む枠組みを提案した。検出精度75%でも平均位置誤差を約50%削減できることを示し、FGSM攻撃のVPRへの有効性も初めて調査した。
- 低コストカメラによる暗闇での長時間露光位置推定位置推定2025/4/1
低コストカメラで撮影した激しくぼやけた暗闇画像を用い、SeqSLAMによる位置推定性能を評価し、夜間に学習した経路を昼間に推定する際の有効性とSeqSLAMの要素の役割を統計的に分析した。
- 動的な水中環境の長期モニタリングのための画像ベースの再定位と位置合わせ水中ロボティクス/Visual Place Recognition2025/3/1
水中映像からVisual Place Recognition・特徴マッチング・セグメンテーションを統合し、再訪領域の同定と生態系変化の解析を可能にするパイプラインを提案。大規模水中VPRベンチマークSQUIDLE+も構築した。
- スパイキングネットワークの閾値適応による最短経路探索と場所の曖昧性解消ニューロモルフィックナビゲーション2025/3/1
スパイキングニューラルネットワークにおいて、スパイクタイミング依存の閾値適応と曖昧性依存の閾値適応を導入し、経路のバックトレースと曖昧な環境での場所特定を可能にすることで、効率的な最短経路探索を実現した研究。
- ロボティクス向け画像ベース地理推定:ブラックボックス視覚言語モデルは実用域に達したかVLA2025/1/1
ブラックボックス設定の生成型視覚言語モデルを単体のゼロショット地理推定システムとして評価し、プロンプトやクエリ画像の変動に対する精度と一貫性を調査した。
- もう正解データはいらない!SfMとVisual SLAMの正解データ不要チューニングSLAM2024/12/1
幾何学的な正解データを使わずに、入力画像へのノイズ付加による感度推定でSfMやVisual SLAMの評価・ハイパーパラメータ調整を可能にする手法を提案した論文。
- 視覚的場所認識における新興トレンドと研究機会の探求VLA2024/11/1
ロボットの自己位置推定やSLAMで重要となる視覚的場所認識について、Vision-Languageモデルなど最近の研究動向と今後の機会を概観したサーベイ。
- マッチドフィルタリングに基づく都市・自然環境向けLiDAR位置認識位置認識2024/9/1
LiDARのBEV記述子とマッチドフィルタリングによる2段階探索で、都市と自然環境の両方で回転・並進不変な位置認識と相対姿勢推定を実現した。
- ROS2WASM: ロボットオペレーティングシステムをウェブへROS/WebAssembly2024/9/1
ROS 2をWebAssemblyにクロスコンパイルし、ブラウザ上でインストール不要に実行できるようにした研究。専用ミドルウェアとWebプラットフォームを開発し、教育や研究の再現性・共有性を高める。
- 超低消費電力オンデバイスロボット自己位置推定のための小型ニューロモーフィックシステムニューロモーフィック/自己位置推定2024/8/1
スパイキングニューラルネットワークとイベントベース視覚センサ、ニューロモーフィックプロセッサを単一チップに統合し、従来の8%未満の消費電力で最大8kmの場所認識を実現した。
- FUSELOC: 大域記述子と局所記述子の融合による視覚位置推定での2D-3Dマッチング曖昧性解消視覚位置推定2024/8/1
局所記述子と大域記述子を重み付き平均で融合し、地理的に近い記述子を特徴空間で近づけることで、直接2D-3Dマッチングの曖昧性を低減し、階層的手法に匹敵する精度を43%少ないメモリと1.6倍の速度で実現した。
- 位置推定の検証によるロボットナビゲーションのための視覚的場所認識の改善VPR/ナビゲーション2024/7/1
視覚的場所認識(VPR)の信頼性を監視するMLPベースの整合性モニタを提案し、実世界実験でナビゲーションと位置推定の精度向上を実証した。
- イベントカメラにおける高速・低速適応バイアス制御による視覚的場所認識の強化視覚的場所認識2024/3/1
イベントカメラのバイアスパラメータをフィードバック制御で自動調整する手法を提案し、視覚的場所認識タスクで評価した。
- Applications of Spiking Neural Networks in Visual Place Recognition2023/11/1
- Collaborative Visual Place Recognition2023/10/1
- VPRTempo: A Fast Temporally Encoded Spiking Neural Network for Visual Place Recognition2023/9/1
- Teach and Repeat Navigation: A Robust Control Approach2023/9/1
- Trajectory Tracking via Multiscale Continuous Attractor Networks2023/8/1
- DisPlacing Objects: Improving Dynamic Vehicle Detection via Visual Place Recognition under Adverse Conditions2023/6/1
- Locking On: Leveraging Dynamic Vehicle-Imposed Motion Constraints to Improve Visual Localization2023/6/1
- A-MuSIC: An Adaptive Ensemble System For Visual Place Recognition In Changing Environments2023/3/1
- Deep Declarative Dynamic Time Warping for End-to-End Learning of Alignment Paths2023/3/1
- Residual Skill Policies: Learning an Adaptable Skill-based Action Space for Reinforcement Learning for Robotics2022/11/1
- Boosting Performance of a Baseline Visual Place Recognition Technique by Predicting the Maximally Complementary Technique2022/10/1
- Improving Worst Case Visual Localization Coverage via Place-specific Sub-selection in Multi-camera Systems2022/6/1
- How Many Events do You Need? Event-based Visual Place Recognition Using Sparse But Varying Pixels2022/6/1
- ReF -- Rotation Equivariant Features for Local Feature Matching2022/3/1
- MultiRes-NetVLAD: Augmenting Place Recognition Training with Low-Resolution Imagery2022/2/1
- PointCrack3D: Crack Detection in Unstructured Environments using a 3D-Point-Cloud-Based Deep Neural Network2021/11/1
- Spiking Neural Networks for Visual Place Recognition via Weighted Neuronal Assignments2021/9/1
- Probabilistic Appearance-Invariant Topometric Localization with New Place Awareness2021/7/1
- A Hierarchical Dual Model of Environment- and Place-Specific Utility for Visual Place Recognition2021/7/1
- Bayesian Controller Fusion: Leveraging Control Priors in Deep Reinforcement Learning for Robotics2021/7/1
- SeqNetVLAD vs PointNetVLAD: Image Sequence vs 3D Point Clouds for Day-Night Place Recognition2021/6/1
- Probabilistic Visual Place Recognition for Hierarchical Localization2021/5/1
- A RoboStack Tutorial: Using the Robot Operating System Alongside the Conda and Jupyter Data Science Ecosystems2021/4/1
- Uncertainty for Identifying Open-Set Errors in Visual Object Detection2021/4/1
- Sequential Place Learning: Heuristic-Free High-Performance Long-Term Place Recognition2021/3/1
- RoRD: Rotation-Robust Descriptors and Orthographic Views for Local Feature Matching2021/3/1
- Where is your place, Visual Place Recognition?2021/3/1
- SeqNet: Learning Descriptors for Sequence-based Hierarchical Place Recognition2021/2/1
- Scene Retrieval for Contextual Visual Mapping2021/2/1
- Semantics for Robotic Mapping, Perception and Interaction: A Survey2021/1/1
- DeepSeqSLAM: A Trainable CNN+RNN for Joint Global Description and Sequence-based Place Recognition2020/11/1
- Fast and Robust Bio-inspired Teach and Repeat Navigation2020/10/1
- Intelligent Reference Curation for Visual Place Recognition via Bayesian Selective Fusion2020/10/1
- Event-based visual place recognition with ensembles of temporal windows2020/6/1
- Robot Perception enables Complex Navigation Behavior via Self-Supervised Learning2020/6/1
- Delta Descriptors: Change-Based Place Representation for Robust Visual Localization2020/6/1
- Multiplicative Controller Fusion: Leveraging Algorithmic Priors for Sample-efficient Reinforcement Learning and Safe Sim-To-Real Transfer2020/3/1
- MVP: Unified Motion and Visual Self-Supervised Learning for Large-Scale Robotic Navigation2020/3/1
- Hierarchical Multi-Process Fusion for Visual Place Recognition2020/2/1
- Fast, Compact and Highly Scalable Visual Place Recognition through Sequence-based Matching of Overloaded Representations2020/1/1
- CityLearn: Diverse Real-World Environments for Sample-Efficient Navigation Policy Learning2019/10/1
- A Hybrid Compact Neural Architecture for Visual Place Recognition2019/10/1
- Residual Reactive Navigation: Combining Classical and Learned Navigation Strategies For Deployment in Unknown Environments2019/9/1
- Automatic Coverage Selection for Surface-Based Visual Localization2019/6/1
- BTEL: A Binary Tree Encoding Approach for Visual Localization2019/6/1
- Filter Early, Match Late: Improving Network-Based Visual Place Recognition2019/6/1
- Multi-Process Fusion: Visual Place Recognition Using Multiple Image Processing Methods2019/3/1
- LookUP: Vision-Only Real-Time Precise Underground Localisation for Autonomous Mining Vehicles2019/3/1
- Look No Deeper: Recognizing Places from Opposing Viewpoints under Varying Scene Appearance using Single-View Depth Estimation2019/2/1
- A Holistic Visual Place Recognition Approach using Lightweight CNNs for Significant ViewPoint and Appearance Changes2018/11/1
- Memorable Maps: A Framework for Re-defining Places in Visual Place Recognition2018/11/1
- Feature Map Filtering: Improving Visual Place Recognition with Convolutional Calibration2018/10/1
- An Orientation Factor for Object-Oriented SLAM2018/9/1
- Learning Deployable Navigation Policies at Kilometer Scale from a Single Traversal2018/7/1
- OpenSeqSLAM2.0: An Open Source Toolbox for Visual Place Recognition Under Changing Conditions2018/4/1
- LoST? Appearance-Invariant Place Recognition for Opposite Viewpoints using Visual Semantics2018/4/1
- QuadricSLAM: Dual Quadrics from Object Detections as Landmarks in Object-oriented SLAM2018/4/1
- The Limits and Potentials of Deep Learning for Robotics2018/4/1
- Don't Look Back: Robustifying Place Categorization for Viewpoint- and Condition-Invariant Place Recognition2018/1/1
- Rhythmic Representations: Learning Periodic Patterns for Scalable Place Recognition at a Sub-Linear Storage Cost2017/12/1
- One-Shot Reinforcement Learning for Robot Navigation with Interactive Replay2017/11/1
- Addressing Challenging Place Recognition Tasks using Generative Adversarial Networks2017/9/1
- Adversarial Discriminative Sim-to-real Transfer of Visuo-motor Policies2017/9/1
- Dual Quadrics from Object Detection BoundingBoxes as Landmark Representations in SLAM2017/8/1
- Deja vu: Scalable Place Recognition Using Mutually Supportive Feature Frequencies2017/7/1
- Look No Further: Adapting the Localization Sensory Window to the Temporal Characteristics of the Environment2017/6/1
- Multi-Modal Trip Hazard Affordance Detection On Construction Sites2017/6/1
- Improving Condition- and Environment-Invariant Place Recognition with Semantic Place Categorization2017/6/1
- Tuning Modular Networks with Weighted Losses for Hand-Eye Coordination2017/5/1
- What Would You Do? Acting by Learning to Predict2017/3/1
- Action Recognition: From Static Datasets to Moving Robots2017/1/1
- Deep Learning Features at Scale for Visual Place Recognition2017/1/1
- 3D tracking of water hazards with polarized stereo cameras2017/1/1
- Modular Deep Q Networks for Sim-to-real Transfer of Visuo-motor Policies2016/10/1
- Meaningful Maps With Object-Oriented Semantic Mapping2016/9/1
- 2D Visual Place Recognition for Domestic Service Robots at Night2016/5/1
- Towards Vision-Based Deep Reinforcement Learning for Robotic Motion Control2015/11/1
- Place Categorization and Semantic Mapping on a Mobile Robot2015/7/1
- Place Recognition with Event-based Cameras and a Neural Implementation of SeqSLAM2015/5/1
- On the Performance of ConvNet Features for Place Recognition2015/1/1