Tao Huang
収録論文 40本 ・ フィジカルAI/ロボット学習
VLA
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- StellaVLA: 文脈構造化デモンストレーションによる汎用視覚言語行動モデルVLA2026/8/12
StellaVLAは、テスト時に単一の構造化デモンストレーションを条件付けすることで、分布外の状況でも適応できる視覚言語行動モデルを提案する。
- StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models2026/8/1
- Revisiting Parameter Redundancy in Vision-Language-Action Models: Insights from VLM-to-VLA Adaptation2026/6/1
- Seeing Realism from Simulation: Efficient Video Transfer for Vision-Language-Action Data Augmentation2026/5/1
- Tactile-based Multimodal Fusion in Embodied Intelligence: A Survey of Vision, Language, and Contact-Driven Paradigms2026/5/1
- SMASH: Mastering Scalable Whole-Body Skills for Humanoid Ping-Pong with Egocentric Vision2026/4/1
- Feel Robot Feels: Tactile Feedback Array Glove for Dexterous Manipulation2026/3/1
- WaterVideoQA: ASV-Centric Perception and Rule-Compliant Reasoning via Multi-Modal Agents2026/2/1
- Affordance Field Intervention: Enabling VLAs to Escape Memory Traps in Robotic Manipulation2025/12/8
- RoboMirror: Understand Before You Imitate for Video to Humanoid Locomotion2025/12/1
- Do You Have Freestyle? Expressive Humanoid Locomotion via Audio Control2025/12/1
- GuangMing-Explorer: A Four-Legged Robot Platform for Autonomous Exploration in General Environments2025/12/1
- Affordance Field Intervention: Enabling VLAs to Escape Memory Traps in Robotic Manipulation2025/12/1
- AerialMind: Towards Referring Multi-Object Tracking in UAV Scenarios2025/11/1
- PhysHSI: Towards a Real-World Generalizable and Natural Humanoid-Scene Interaction System2025/10/1
- From Language to Locomotion: Retargeting-free Humanoid Control via Motion Latent Guidance2025/10/1
- Towards Adaptable Humanoid Control via Adaptive Motion Tracking2025/10/1
- Humanoid Goalkeeper: Learning from Position Conditioned Task-Motion Constraints2025/10/1
- GLEAM: Learning Generalizable Exploration Policy for Active Mapping in Complex 3D Indoor Scenes2025/5/1
- Unsupervised Radar Point Cloud Enhancement via Arbitrary LiDAR Guided Diffusion Prior2025/5/1
- OptiPMB: Enhancing 3D Multi-Object Tracking with Optimized Poisson Multi-Bernoulli Filtering2025/3/1
- BeamDojo: Learning Agile Humanoid Locomotion on Sparse Footholds2025/2/14
- Learning Humanoid Standing-up Control across Diverse Postures2025/2/12
- VLA-Cache: Efficient Vision-Language-Action Manipulation via Adaptive Token Caching2025/2/4
- BeamDojo: Learning Agile Humanoid Locomotion on Sparse Footholds2025/2/1
- Learning Humanoid Standing-up Control across Diverse Postures2025/2/1
- VB-Com: Learning Vision-Blind Composite Humanoid Locomotion Against Deficient Perception2025/2/1
- VLA-Cache: Efficient Vision-Language-Action Manipulation via Adaptive Token Caching2025/2/1
- Learning Humanoid Locomotion with Perceptive Internal Model2024/11/1
- Robots Pre-train Robots: Manipulation-Centric Robotic Representation from Large-Scale Robot Datasets2024/10/1
- Multi-robot Task Allocation and Path Planning with Maximum Range Constraints2024/9/1
- GRUtopia: Dream General Robots in a City at Scale2024/7/1
- LiDAR Point Cloud-based Multiple Vehicle Tracking with Probabilistic Measurement-Region Association2024/3/1
- Diffusion Reward: Learning Rewards via Conditional Video Diffusion2023/12/1
- Value-Informed Skill Chaining for Policy Learning of Long-Horizon Tasks with Surgical Robot2023/7/1
- Homography matrix based trajectory planning method for robot uncalibrated visual servoing2023/3/1
- Demonstration-Guided Reinforcement Learning with Efficient Exploration for Task Automation of Surgical Robot2023/2/1
- Human-in-the-loop Embodied Intelligence with Interactive Simulation Environment for Surgical Robot Learning2023/1/1
- Cooperative trajectory planning algorithm of USV-UAV with hull dynamic constraints2022/9/1
- Efficient Trajectory Planning and Control for USV with Vessel Dynamics and Differential Flatness2022/9/1