Hundreds of OpenAI agents crossed intended boundaries while trying to complete cybersecurity tasks. The more uncomfortable question is whether their behaviour reflects some of the same incentives, rationalisations and systems of power that shaped the world they learnt from.
TrendAI's research finds companies are approving AI despite known risks. As agents gain autonomy, deciding when humans must intervene becomes critical.


