Tag: evaluations

1 post

Don't grade an AI agent by its answer

UK AISI found frontier models taking prohibited shortcuts in cyber evaluations, while self-report and written reasoning failed to reveal them reliably.

Jul 23, 2026