Beyond Task Reward: A Controller-Restriction Protocol for Evaluating Embodiment-Dependent Competence
Organizations: Robotics Department, University of Michigan, Ann Arbor, MI, USA
Abstract
Co-design methods optimize a robot's body and controller jointly and judge the result by one number, the task reward of the fully optimized pair. That number cannot separate morphologies whose competence depends on the controller to very different degrees. We evaluate a morphology by restricting its controller instead, recording the task competence it retains under an explicitly declared, low-complexity controller family, environment, task and search budget. On three EvoGym locomotion tasks, task reward explains only , and of the variance in this quantity, and geometric descriptors do not predict it under run-grouped cross-validation. The measurement is reliable across optimizer restarts (ICC--) but depends on the declared family: phasing the drive by actuator index instead of position ranks the same morphologies at Spearman -- and reverses reward-matched pairs. As a second search objective the axis improved competence at matched task reward in of paired runs, short of a pre-registered bar of . Used after an ordinary reward-only search instead, to choose within its top task-reward band, it selected a different body in all runs offering a choice, at a cost of at most reward units, and in of that body also scored higher under a held-out family. Restricted-control competence is therefore a reportable property of a co-designed morphology, interpretable only with the controller family that defines it.
Figures & tables
| Seed | Thr. | MO | Reward | |||
|---|---|---|---|---|---|---|
| — | empty | |||||
| — | empty | |||||
| Run | C2 | C3 | C3 ′ | Cost | |
|---|---|---|---|---|---|
| D, s0 † | |||||
| D, s1 † | |||||
| D, s2 † | |||||
| D, s3 † | |||||
| D, s1 | |||||
| W, s2 |