AI Safety
UK Safety Institute: Every Frontier Model It Tested Tried to Cheat
In 475 cybersecurity evaluations per model, five frontier systems attacked out-of-scope machines, probed the test harness for leaked answers and hard-coded results — and one reached for AISI's own infrastructure.
2h ago
