| 1 |
Pointing-VLA: Typed Spatial Grounding Interfaces for Vision-Language-Action Manipulation |
提出Pointing-VLA以解决视觉-语言-动作中的空间定位问题 |
manipulation vision-language-action VLA |
|
|
| 2 |
InstructMove: A Text-Indispensable Benchmark for Instruction-Following Manipulation |
提出InstructMove以解决指令跟随操作的评估问题 |
manipulation physically plausible vision-language-action |
✅ |
|
| 3 |
Think Only When Needed: Prompt-Authority Control for Selective Slow-Path Intervention in Vision-Language-Action Manipulation |
提出TOWN-VLA以解决视觉-语言-动作操控中的干预问题 |
manipulation vision-language-action VLA |
|
|
| 4 |
Guided Riemannian Optimization (GuRO): Bridging Model Predictive Control and Decision Transformers |
提出GuRO框架以解决高维非线性系统决策问题 |
quadruped MPC model predictive control |
|
|
| 5 |
Triplet2Track: A Hierarchical System with Object-Centric Representations for Reliable Long-Horizon Manipulation |
提出Triplet2Track系统以解决长时间操控中的可靠性问题 |
manipulation imitation learning VLA |
|
|
| 6 |
Design of a Biomimetic Joint-Covering Skin with Tissue-Like Structure to Enhance Proprioception in a Musculoskeletal Humanoid |
提出生物仿生关节覆盖皮肤以增强人形机器人本体感觉 |
humanoid |
|
|