arXiv Artificial Intelligence

RobotValues: Evaluating Household Robots When Human Values Conflict

RobotValues: Evaluating Household Robots When Human Values Conflict

Quick summary

arXiv:2606.03312v2 Announce Type: replace-cross Abstract: While household robots are often evaluated based on task completion, everyday domestic environments involve value-conflicting situations where robots are expected to choose actions that prioritize diverse values such as human autonomy, efficiency, or social appropriateness. Yet, there are no benchmarks for evaluating robots' value preferences in such scenarios. We introduce RobotValues, a benchmark to evaluate household robot planners in 8K value-conflict scenarios. Each instance consists of a realistic, synthetically generated househol

Key takeaways

  • arXiv:2606.03312v2 Announce Type: replace-cross Abstract: While household robots are often evaluated based on task completion, everyday domestic environments involve value-conflicting situations where robots are expected to choose actions that prioritize diverse values such as human autonomy, efficiency, or social appropriateness.
  • Yet, there are no benchmarks for evaluating robots' value preferences in such scenarios.
  • We introduce RobotValues, a benchmark to evaluate household robot planners in 8K value-conflict scenarios.

Why it matters

“RobotValues: Evaluating Household Robots When Human Values Conflict” highlights the need for repeatable measurement rather than a single impressive demonstration. Independent validation across datasets and clearly stated limitations determine whether a result can guide product decisions.

Kaynak sitede devamını oku: arXiv Artificial Intelligence ↗