Security
AI-enabled cyber risk, model misuse, safeguards, containment, access controls and incident response.
-
Agent containment is becoming an ordinary security problem
Recent evaluations show that capable agents can find routes outside the boundaries researchers intended to give them.
-
The safety problem is getting faster
AI is increasingly helping build and operate AI. The important safety questions are moving from theory into engineering.