AIIC AI Intelligence Centre

SOURCE-LINKED INTELLIGENCE

One Word, Different Action: A Real-Robot Benchmark for Language-Conditioned Embodied Reasoning

arXiv · AI, language, vision and robotics · article · Sep 4, 2026 · UTC

Changes in natural-language instructions can directly alter the behavior ultimately executed by a robot, but such changes do not always imply that the task itself has changed. We propose One Word, Different Action, a language-conditioned executable decision benchmark based on real-robot physical decision anchors. It evaluates robot behavioral responses under task-preserving and task-changing conditions and examines joint reasoning over multiple task constraints. The benchmark uses decision invariance and decision sensitivity to measure action stability when task semantics are preserved and cor

Read original source ↗ Open in workspace

recordType
paper
region
Global

Evidence & attribution

First collected: 2026-09-20T21:52:07.471Z. This is not the publication date.