Wenlong Huang
収録論文 18本 ・ フィジカルAI/ロボット学習
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- Masked Visual Actions for Unified World Modeling2026/7/1
- CoStream: Composing Simple Behaviors for Generalizable Complex Manipulation2026/6/1
- SaiVLA-0: Cerebrum--Pons--Cerebellum Tripartite Architecture for Compute-Aware Vision-Language-Action2026/3/9
- SaiVLA-0: Cerebrum--Pons--Cerebellum Tripartite Architecture for Compute-Aware Vision-Language-Action2026/3/1
- ROI-Driven Foveated Attention for Unified Egocentric Representations in Vision-Language-Action Systems2026/3/1
- PointWorld: Scaling 3D World Models for In-The-Wild Robotic Manipulation2026/1/1
- Dream2Flow: Bridging Video Generation and Open-World Manipulation with 3D Object Flow2025/12/1
- ENACT: Evaluating Embodied Cognition with World Modeling of Egocentric Interaction2025/11/1
- UAD: Unsupervised Affordance Distillation for Generalization in Robotic Manipulation2025/6/1
- A Real-to-Sim-to-Real Approach to Robotic Manipulation with VLM-Generated Iterative Keypoint Rewards2025/2/1
- ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation2024/9/1
- VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models2023/7/1
- PaLM-E: An Embodied Multimodal Language Model2023/3/1
- Grounded Decoding: Guiding Text Generation with Grounded Models for Embodied Agents2023/3/1
- Code as Policies: Language Model Programs for Embodied Control2022/9/1
- Inner Monologue: Embodied Reasoning through Planning with Language Models2022/7/1
- Language Models as Zero-Shot Planners: Extracting Actionable Knowledge for Embodied Agents2022/1/1
- Generalization in Dexterous Manipulation via Geometry-Aware Multi-Task Learning2021/11/1