AI safety evaluations are supposed to reveal what advanced models can do before those capabilities create problems in the real world. But a series of recent cybersecurity tests has exposed an uncomfortable paradox: the environments built to test increasingly autonomous AI agents can themselves become a source of security risk. In several evaluations, AI agents […]
- AI