arXiv Artificial Intelligence

Beyond Task Completion: Training Capable and Safe Computer-Use Agents

Beyond Task Completion: Training Capable and Safe Computer-Use Agents

Quick summary

arXiv:2609.22178v2 Announce Type: cross Abstract: Computer-use agents (CUAs) have made rapid progress in completing complex tasks through graphical user interfaces, yet post-training centered on task success alone does not induce reliable safety behavior. A reliable CUA must condition its execution on risk: it should complete ordinary benign tasks, avoid environmental hazards and continue when a safe completion path remains, and refuse when the goal is harmful or no safe path exists. To learn this conditional policy, we develop Safety and Capability Optimization for Policy Execution (SCOPE), w

Key takeaways

  • arXiv:2609.22178v2 Announce Type: cross Abstract: Computer-use agents (CUAs) have made rapid progress in completing complex tasks through graphical user interfaces, yet post-training centered on task success alone does not induce reliable safety behavior.
  • A reliable CUA must condition its execution on risk: it should complete ordinary benign tasks, avoid environmental hazards and continue when a safe completion path remains, and refuse when the goal is harmful or no safe path exists.
  • To learn this conditional policy, we develop Safety and Capability Optimization for Policy Execution (SCOPE), w

Why it matters

The significance is not only the legal text but how it changes product design. Decisions around “Beyond Task Completion: Training Capable and Safe Computer-Use Agents” may reshape data collection, model training, output accountability and market access.

Kaynak sitede devamını oku: arXiv Artificial Intelligence ↗