SOURCE-LINKED INTELLIGENCE
One Word, Different Action: A Real-Robot Benchmark for Language-Conditioned Embodied Reasoning
Changes in natural-language instructions can directly alter the behavior ultimately executed by a robot, but such changes do not always imply that the task itself has changed. We propose One Word, Different Action, a language-conditioned executable decision benchmark based on real-robot physical decision anchors. It evaluates robot behavioral responses under task-preserving and task-changing conditions and examines joint reasoning over multiple task constraints. The benchmark uses decision invariance and decision sensitivity to measure action stability when task semantics are preserved and cor
Read original source ↗ Open in workspace
- recordType
- paper
- region
- Global
Evidence & attribution
- arXiv · AI, language, vision and robotics · 2026-09-04T15:21:58.000Z
First collected: 2026-09-20T21:52:07.471Z. This is not the publication date.