Definition

Agent safety covers practices, controls, and governance for deploying AI agents that can take autonomous actions — browsing, editing, calling APIs, or modifying external systems — without causing unintended harm to third-party infrastructure or data.

Key Points

Sources