Hanqing Wang
Shanghai Artificial Intelligence Laboratory
収録論文 30本 ・ フィジカルAI/ロボット学習
シミュレーション評価
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- GAUGE: 物理的忠実性を測定するための実世界基盤ベンチマークシミュレーション評価2026/8/6
シミュレーションエンジンと生成ビデオワールドモデルの物理的忠実性を、実世界の軌跡に基づいて診断するベンチマークを提案した。
- GAUGE: A Measurement-Grounded Benchmark for Physical Fidelity in Simulation Engines and Video World Models2026/8/1
- InternVLA-A1.5: Unifying Understanding, Latent Foresight, and Action for Compositional Generalization2026/7/1
- RoboSnap: One-Shot Real-to-Sim Scene Generation for Generalizable Robot Learning and Evaluation2026/7/1
- Exploratory, Communicative, and Deployable: Vision-Driven Embodied Agents for Open-World Mobile Manipulation2026/7/1
- EBench: Elemental Diagnosis of Generalist Mobile Manipulation Policies2026/6/1
- Perfect Demo Makes Poor Teacher: Learning Robust Alignment from Critical Motion Segments2026/6/1
- Affordance2Action: Task-Conditioned Scene-level Affordance Grounding for Real-Time Manipulation2026/6/1
- Event-VLA: Action-Conditioned Event Fusion for Robust Vision-Language-Action Model2026/6/1
- SIM1: Physics-Aligned Simulator as Zero-Shot Data Scaler in Deformable Worlds2026/4/1
- Tac2Real: Reliable and GPU Visuotactile Simulation for Online Reinforcement Learning and Zero-Shot Real-World Deployment2026/3/1
- FSAG: Enhancing Human-to-Dexterous-Hand Finger-Specific Affordance Grounding via Diffusion Models2026/1/1
- InternVLA-A1: Unifying Understanding, Generation and Action for Robotic Manipulation2026/1/1
- A4-Agent: An Agentic Framework for Zero-Shot Affordance Reasoning2025/12/1
- VL-LN Bench: Towards Long-horizon Goal-oriented Navigation with Active Dialogs2025/12/1
- VLNVerse: A Benchmark for Vision-Language Navigation with Versatile, Embodied, Realistic Simulation and Evaluation2025/12/1
- InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy2025/10/1
- InternScenes: A Large-scale Simulatable Indoor Scene Dataset with Realistic Layouts2025/9/1
- Affordance-R1: Reinforcement Learning for Generalizable Affordance Reasoning in Multimodal Large Language Model2025/8/1
- StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling2025/7/1
- InstructVLA: Vision-Language-Action Instruction Tuning from Understanding to Manipulation2025/7/1
- Rethinking the Embodied Gap in Vision-and-Language Navigation: A Holistic Study of Physical and Visual Disparities2025/7/1
- GENMANIP: LLM-driven Simulation for Generalizable Instruction-Following Manipulation2025/6/1
- CronusVLA: Towards Efficient and Robust Manipulation via Multi-Frame Vision-Language-Action Modeling2025/6/1
- NavDP: Learning Sim-to-Real Navigation Diffusion Policy with Privileged Information Guidance2025/5/1
- TeleOpBench: A Simulator-Centric Benchmark for Dual-Arm Dexterous Teleoperation2025/5/1
- LabUtopia: High-Fidelity Simulation and Hierarchical Benchmark for Scientific Embodied Agents2025/5/1
- Open-Vocabulary Object-Goal Navigation by Generalizing Semantic Mapping with Dense CLIP2024/7/1
- GRUtopia: Dream General Robots in a City at Scale2024/7/1
- ETPNav: Evolving Topological Planning for Vision-Language Navigation in Continuous Environments2023/4/1