Breaking the Vision-Action Shortcut: Latent Interface Training for Generalizable Robotics Foundation Models Paper • 2609.12641 • Published 5 days ago • 70
RL^2-VLA: Adaptive RL Latent Compositional Steering with Test-Time Scaling for Vision-Language-Action Models Paper • 2607.26991 • Published Jul 30 • 10
RL^2-VLA: Adaptive RL Latent Compositional Steering with Test-Time Scaling for Vision-Language-Action Models Paper • 2607.26991 • Published Jul 30 • 10
RL$^2$-VLA: Adaptive RL Latent Compositional Steering with Test-Time Scaling for Vision-Language-Action Models Paper • 2607.26991 • Published Jul 30 • 10
RL$^2$-VLA: Adaptive RL Latent Compositional Steering with Test-Time Scaling for Vision-Language-Action Models Paper • 2607.26991 • Published Jul 30 • 10 • 3
Search-TTA-RL-VLN Collection Test-Time Adaptation Framework for Multimodal Visual Navigation and Search (https://search-tta.github.io/) • 7 items • Updated Mar 2