Jingjing Gong
Shanghai Innovation Institute
収録論文 20本 ・ フィジカルAI/ロボット学習
エージェントアーキテクチャ
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- ETA: 身体性タスクのための新しいエージェントパラダイムエージェントアーキテクチャ2026/8/4
ロボットのChatGPT的瞬間を目指し、プランナー・インターフェース・ワールドのループで構成される新しい身体性タスクエージェント(ETA)を提案し、オープンソース実装OpenETAを公開した。
- ETA: A New Agentic Paradigm for Embodied Tasks2026/8/1
- WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning2026/7/1
- Learning to Move Before Learning to Do: Task-Agnostic pretraining for VLAs2026/7/1
- HiMe: Hierarchical Embodied Memory for Long-Horizon Vision-Language-Action Control2026/7/1
- CoRE-VLA: Towards Scalable and Robust Vision-Language-Action Modeling via Conditional Routing of Experts2026/7/1
- Let It Be Simple: One-Step Action Generation for Vision-Language-Action Models2026/6/4
- Two Bridges, One Pathway: From VLMs to Generalizable VLAs with Embodied Trajectory-Coupled Data2026/6/1
- Let It Be Simple: One-Step Action Generation for Vision-Language-Action Models2026/6/1
- In-Context World Modeling for Robotic Control2026/6/1
- Advancing Omnimodal Embodied Agents from Isolated Skills to Everyday Physical Autonomy2026/6/1
- Coarse-to-Control: Action-Token Planning for Vision-Language-Action Models2026/6/1
- World Action Models: The Next Frontier in Embodied AI2026/5/1
- ActionCodec: What Makes for Good Action Tokenizers2026/2/1
- FRoM-W1: Towards General Humanoid Whole-Body Control with Language Instructions2026/1/1
- FASTer: Toward Efficient Autoregressive Vision Language Action Modeling via Neural Action Tokenization2025/12/1
- SRPO: Self-Referential Policy Optimization for Vision-Language-Action Models2025/11/1
- RoboOmni: Proactive Robot Manipulation in Omni-modal Context2025/10/1
- LIBERO-Plus: In-depth Robustness Analysis of Vision-Language-Action Models2025/10/1
- World-aware Planning Narratives Enhance Large Vision-Language Model Planner2025/6/1