站内快照 · 国内可打开。外网原文可能无法访问。
- 论文公开站arXiv
VersaCamVLA: Camera-Configurable VLA Policies for Robotic Manipulation
Vision-Language-Action (VLA) models have emerged as powerful foundations for robotic manipulation, but their reliance on fixed camera configurations during training makes them brittle to changes in camera count or pose during deployment. To overcome these limitations, we propose