Co-VLA: Coordination-Aware Structured Action Modeling for Dual-Arm Vision-Language-Action Systems
FeaturedYandong Wang, Jiaqian Yu, Xiongfeng Peng, Lu Xu, Yamin Mao, Weiming Li, Jaewook Yoo, Dongwook Lee, Daehyun Ji, Mingbo Zhao, Chao Zhang · · 2026
Framework
N/A
License
N/A
Stars
N/A
Summary
Vision-language-action (VLA) models show strong capabilities in single and dual-arm robotic manipulation. Prior works show coordinated bimanual behaviors can emerge from end-to-end...
Abstract Summary
Key Points
- Proposes Co-VLA for robotics.
- Evaluated on real or simulated robotic tasks.
- Primary contribution in ar-vr.
Related Papers
Co-VLA: Coordination-Aware Structured Action Modeling for Dual-Arm Vision-Language-Action Systems
Yandong Wang, Jiaqian Yu, Xiongfeng Peng et al. · arXiv preprint · Jun 2026
Vision-language-action (VLA) models show strong capabilities in single and dual-arm robotic manipulation. Prior works show coordinated bimanual behaviors can emerge from end-to-end...
Invertible Neural Network Adapter for One-Step Flow Matching in Robot Manipulation
Yu Zhang, Kangyi Ji, Yongxiang Zou et al. · arXiv preprint · Jun 2026
This paper presents an invertible neural network adapter for general robotic manipulation, designed to generate precise high-dimensional actions conditioned on multimodal observati...
Invertible Neural Network Adapter for One-Step Flow Matching in Robot Manipulation
Yu Zhang, Kangyi Ji, Yongxiang Zou et al. · arXiv preprint · Jun 2026
This paper presents an invertible neural network adapter for general robotic manipulation, designed to generate precise high-dimensional actions conditioned on multimodal observati...
Neuro-Symbolic Safety Guidance for Vision-Language-Action Models via Constrained Flow Matching
William English, Hao Zheng, Rickard Ewetz · arXiv preprint · Jul 2026
Vision-Language-Action (VLA) models have demonstrated promising generalization capabilities across robotic manipulation tasks, yet their real-world deployment remains limited by th...