Yu Wang
収録論文 83本 ・ フィジカルAI/ロボット学習
強化学習/UAV
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- AgilePE: 自己対戦強化学習による自律UAV追跡・回避強化学習/UAV2026/8/14
自己対戦強化学習を用いて、UAVの追跡・回避行動をエンドツーエンドで学習し、シミュレーションから実機へのゼロショット転送を実現したシステムを提案する。
- Optimal Constrained sc-LTL Planning in MDPs via Switching Policies2026/8/1
- Exact Model-Free Policy Iteration for Co-safe LTL Planning2026/8/1
- Harness VLA: Steering Frozen VLAs into Reliable Manipulation Primitives via Memory-Guided Agents2026/7/1
- The Curse of Precision: A Data Scaling Law for High-Precision Robotic Manipulation2026/7/1
- Cosmos 3: Omnimodal World Models for Physical AI2026/6/1
- Motion-Focused Latent Action Enables Cross-Embodiment VLA Training from Human EgoVideos2026/6/1
- Bionic Human-Motion Style Transfer for Physically Executable Whole-Body Control of Humanoid Robots2026/6/1
- STEAM: Self-Supervised Temporal Ensemble Advantage Modeling for Real-World Robot Learning2026/6/1
- Human2Humanoid: Physics-Aware Cross-Morphology Motion Retargeting for Humanoid Robots2026/6/1
- StreamingVLA: Streaming Vision-Language-Action Model with Action Flow Matching and Adaptive Early Observation2026/3/1
- AsgardBench -- Evaluating Visually Grounded Interactive Planning Under Minimal Feedback2026/3/1
- RLinf-USER: A Unified and Extensible System for Real-World Online Policy Learning in Embodied AI2026/2/1
- WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL2026/2/1
- Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models2026/2/1
- Mind the Gap: Learning Implicit Impedance in Visuomotor Policies via Intent-Execution Mismatch2026/2/1
- SKATER: Synthesized Kinematics for Advanced Traversing Efficiency on a Humanoid Robot via Roller Skate Swizzles2026/1/1
- ArtiSG: Functional 3D Scene Graph Construction via Human-demonstrated Articulated Objects Manipulation2025/12/1
- $π_\texttt{RL}$: Online RL Fine-tuning for Flow-based Vision-Language-Action Models2025/10/29
- RLinf-VLA: A Unified and Efficient Framework for Reinforcement Learning of Vision-Language-Action Models2025/10/1
- JuggleRL: Mastering Ball Juggling with a Quadrotor via Deep Reinforcement Learning2025/9/1
- SAC Flow: Sample-Efficient Reinforcement Learning of Flow-Based Policies via Velocity-Reparameterized Sequential Modeling2025/9/1
- SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning2025/9/1
- Traversing Narrow Paths: A Two-Stage Reinforcement Learning Framework for Robust and Safe Humanoid Walking2025/8/1
- D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning2025/8/1
- Control Synthesis in Partially Observable Environments for Complex Perception-Related Objectives2025/7/1
- Lasso Gripper: A String Shooting-Retracting Mechanism for Shape-Adaptive Grasping2025/6/1
- ReinFlow: Fine-tuning Flow Matching Policy with Online Reinforcement Learning2025/5/1
- Hysteresis-Aware Neural Network Modeling and Whole-Body Reinforcement Learning Control of Soft Robots2025/4/1
- Multi-Robot System for Cooperative Exploration in Unknown Environments: A Survey2025/3/1
- High-Precision Transformer-Based Visual Servoing for Humanoid Robots in Aligning Tiny Objects2025/3/1
- HEATS: A Hierarchical Framework for Efficient Autonomous Target Search with Mobile Manipulators2025/3/1
- Real-Time LiDAR Point Cloud Compression and Transmission for Resource-constrained Robots2025/2/1
- Learning from Suboptimal Data in Continuous Control via Auto-Regressive Soft Q-Network2025/2/1
- VolleyBots: A Testbed for Multi-Drone Volleyball Game Combining Motion Control and Strategic Play2025/2/1
- What Matters in Learning A Zero-Shot Sim-to-Real RL Policy for Quadrotor Control? A Comprehensive Study2024/12/1
- MR-COGraphs: Communication-efficient Multi-Robot Open-vocabulary Mapping System via 3D Scene Graphs2024/12/1
- Neural Internal Model Control: Learning a Robust Control Policy via Predictive Error Feedback2024/11/1
- SPF-EMPC Planner: A real-time multi-robot trajectory planner for complex environments with uncertainties2024/10/1
- Multi-UAV Formation Control with Static and Dynamic Obstacle Avoidance via Reinforcement Learning2024/10/1
- Behavior evolution-inspired approach to walking gait reinforcement training for quadruped robots2024/9/1
- Online Planning for Multi-UAV Pursuit-Evasion in Unknown Environments Using Deep Reinforcement Learning2024/9/1
- Human-Robot Cooperative Distribution Coupling for Hamiltonian-Constrained Social Navigation2024/9/1
- Convergence Guarantee of Dynamic Programming for LTL Surrogate Reward2024/8/1
- FlightBench: Benchmarking Learning-based Methods for Ego-vision-based Quadrotors Navigation2024/6/1
- On the Uniqueness of Solution for the Bellman Equation of LTL Objectives2024/4/1
- Localization matters too: How localization error affects UAV flight2024/3/1
- MASP: Scalable GNN-based Planning for Multi-Agent Navigation2023/12/1
- EVI-SAM: Robust, Real-time, Tightly-coupled Event-Visual-Inertial State Estimation and 3D Dense Mapping2023/12/1
- Active Neural Topological Mapping for Multi-Agent Exploration2023/11/1
- Large Trajectory Models are Scalable Motion Predictors and Planners2023/10/1
- Adaptive Tuning of Robotic Polishing Skills based on Force Feedback Model2023/10/1
- OmniDrones: An Efficient and Flexible Platform for Reinforcement Learning in Drone Control2023/9/1
- ClusterFusion: Real-time Relative Positioning and Dense Reconstruction for UAV Cluster2023/4/1
- HybridFusion: LiDAR and Vision Cross-Source Point Cloud Fusion2023/4/1
- Learning Graph-Enhanced Commander-Executor for Multi-Agent Navigation2023/2/1
- Asynchronous Multi-Agent Reinforcement Learning for Efficient Real-Time Multi-Robot Cooperative Exploration2023/1/1
- Edge-based Monocular Thermal-Inertial Odometry in Visually Degraded Environments2022/10/1
- Point Cloud Change Detection With Stereo V-SLAM:Dataset, Metrics and Baseline2022/7/1
- Robotic Computing on FPGAs: Current Progress, Research Challenges, and Opportunities2022/5/1
- Explore-Bench: Data Sets, Metrics and Evaluations for Frontier-based and Deep-reinforcement-learning-based Autonomous Exploration2022/2/1
- Multi-UAV Coverage Planning with Limited Endurance in Disaster Environment2022/1/1
- Multi-Agent Vulnerability Discovery for Autonomous Driving with Hazard Arbitration Reward2021/12/1
- Relative Distributed Formation and Obstacle Avoidance with Multi-agent Reinforcement Learning2021/11/1
- A drl based distributed formation control scheme with stream based collision avoidance2021/9/1
- An Internal Arc Fixation Channel and Automatic Planning Algorithm for Pelvic Fracture2021/7/1
- Tactile Sensing with a Tendon-Driven Soft Robotic Finger2021/7/1
- The Grasps Under Varied Object Orientation Dataset: Relation Between Grasps and Object Orientation2021/6/1
- Reinforcement Learning with Temporal Logic Constraints for Partially-Observable Markov Decision Processes2021/4/1
- Model-Free Learning of Safe yet Effective Controllers2021/3/1
- Learning Optimal Strategies for Temporal Tasks in Stochastic Games2021/2/1
- Secure Planning Against Stealthy Attacks via Model-Free Reinforcement Learning2020/11/1
- Attentional Separation-and-Aggregation Network for Self-supervised Depth-Pose Learning in Dynamic Scenes2020/11/1
- A Learning-Based Tune-Free Control Framework for Large Scale Autonomous Driving System Deployment2020/11/1
- DRF: A Framework for High-Accuracy Autonomous Driving Vehicle Modeling2020/11/1
- Model-Free Reinforcement Learning for Stochastic Games with Linear Temporal Logic Objectives2020/10/1
- DL-IAPS and PJSO: A Path/Speed Decoupled Trajectory Optimization and its Application in Autonomous Driving2020/9/1
- A Survey of FPGA-Based Robotic Computing2020/9/1
- TDR-OBCA: A Reliable Planner for Autonomous Driving in Free-Space Environment2020/9/1
- Hyperproperties for Robotics: Planning via HyperLTL2019/11/1
- Review of Learning-based Longitudinal Motion Planning for Autonomous Vehicles: Research Gaps between Self-driving and Traffic Congestion2019/10/1
- Control Synthesis from Linear Temporal Logic Specifications using Model-Free Reinforcement Learning2019/9/1
- Recognition of Pyralidae Insects Using Intelligent Monitoring Autonomous Robot Vehicle in Natural Farm Scene2019/3/1