The agent followed the instructions. The system still failed.
Recent frontier-model incidents show why agent safety cannot depend on the prompt alone. Businesses need enforceable boundaries, least privilege, monitored actions, approval gates, and a tested stop path.
Read the field note