返回动态
Kaixi Feng, Guoheng Sun, Ang li

PAIR: Bridging Perception and Action in Vision-Language-Action Models

arXiv:2610.09016v1 Announce Type: new Abstract: Vision-language-action (VLA) models map visual observations and language instructions to continuous robot actions. This task requires a transition from representations that describe the scene and instruction to representations that support action generation. Many continuous-action VLAs leave this transition implicit and supervise it mainly through the final action-predi…