Products & tools · first seen 2h ago, updated 2h ago
How do you safely test an AI agent that’s trying to break things?
OpenAI’s rogue-agent incidents show the trade-off at the heart of cyber evals: Giving models the tools they need to prove themselves can also give them a way out.
Summary from Fast Company - AI.
Coverage 1 article · 1 outlet
-
Fast Company - AIHow do you safely test an AI agent that’s trying to break things?